Synth-JDoc: Synthesizing a Japanese Document Image Dataset for OCR with Diverse Layouts and Embedded Images
Evaluation results demonstrate that the synthetic dataset constructed is the most effective approach for improving LVLM performance on reading vertically written Japanese text.
Keito Sasagawa, Shuhei Kurita, Daisuke Kawahara
· IEEE International Conferenc... · 0 citations