Preprint
Jul 2026
OvisOCR2 Technical Report
This work introduces OvisOCR2, a 0.8B document parsing model designed as an end-to-end parser that combines filtered real-document annotations with synthetic pages whose rendered images and Markdown targets are derived from the same HTML source.
Shiyin Lu, Yinglun Li, Yu Xia et al.
· 1 citation