9 items with this tag.2026年8月06日Deformable DETR: Deformable Transformers for End-to-End Object Detectiondetectiondetrdeformable-attentionmulti-scalefrontendmilestone2026年8月06日End-to-End Object Detection with Transformers (DETR)detectiondetrtransformerset-predictionmilestonefrontend2026年8月06日DINO: DETR with Improved DeNoising Anchor Boxesdetectiondetrdenoisingquery-selectionsotafrontendmilestone2026年7月25日dots.ocr: Multilingual Document Layout Parsing (1.2B vision encoder + 1.7B decoder, ~3B)ocrdocument-parsingunified-vlmlayoutmultilingualend-to-endfrontendbackend2026年7月23日Dolphin-v2: Scalable Anchor Prompting for Document Parsingocrdocument-parsinganchor-promptingphotographed-docsfine-grained-detectionhybrid-parsingfrontend2026年7月23日Dolphin: Document Image Parsing via Heterogeneous Anchor Promptingocrdocument-parsinganchor-promptingparallel-decodingtwo-stageanalyze-then-parsefrontendbackend2026年7月23日MinerU2.5: A Decoupled VLM for Efficient High-Resolution Document Parsingocrdocument-parsingvlmdecoupledhigh-resolutioncoarse-to-finefrontendbackend2026年7月23日PaddleOCR-VL: 0.9B Ultra-Compact VLM for Document Parsingocrdocument-parsingvlmsmall-modellayoutomnidocbenchfrontendbackend2026年7月23日Youtu-Parsing: High-Parallelism Decoding for Document Parsingocrdocument-parsingparallel-decodingtoken-parallelismquery-parallelismregion-promptbackendfrontend