4 items with this tag.2026年7月25日dots.ocr: Multilingual Document Layout Parsing (1.2B vision encoder + 1.7B decoder, ~3B)ocrdocument-parsingunified-vlmlayoutmultilingualend-to-endfrontendbackend2026年7月23日LayoutLMv3: Pre-training for Document AI with Unified Text and Image Maskingocrdocument-understandinglayoutpretrainingmilestone2026年7月23日Logics-Parsing: End-to-End LVLM with RL for Layout & Reading Orderocrdocument-parsingreinforcement-learningreading-orderlayouthtml-outputend-to-end2026年7月23日PaddleOCR-VL: 0.9B Ultra-Compact VLM for Document Parsingocrdocument-parsingvlmsmall-modellayoutomnidocbenchfrontendbackend