8 items with this tag.2026年7月25日dots.ocr: Multilingual Document Layout Parsing (1.2B vision encoder + 1.7B decoder, ~3B)ocrdocument-parsingunified-vlmlayoutmultilingualend-to-endfrontendbackend2026年7月25日Ovis-OCR2: 0.8B End-to-End Document Parsing (OmniDocBench SOTA)ocrdocument-parsingend-to-endsmall-modeldata-enginesynthetic-datadistillationmilestone2026年7月23日General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model (GOT)ocrocr-20unifiedend-to-endmilestone2026年7月23日HunyuanOCR: Commercial-grade Lightweight 1B OCR VLMocrvlmsmall-modelunifiedspottingtranslationend-to-end2026年7月23日Logics-Parsing: End-to-End LVLM with RL for Layout & Reading Orderocrdocument-parsingreinforcement-learningreading-orderlayouthtml-outputend-to-end2026年7月23日OCRVerse: Towards Holistic OCR in End-to-Endocrholistic-ocrend-to-enddata-engineeringsft-rltext-centricvision-centric2026年7月23日POINTS-Reader: Distillation-Free Adaptation of VLMs for Document Conversionocrdocument-parsingdistillation-freesynthetic-dataself-improvementend-to-end2026年7月23日Qianfan-OCR: Unified End-to-End Document Intelligence (Layout-as-Thought)ocrdocument-intelligenceend-to-endlayout-as-thoughtthinkingreading-orderomnidocbench