PP-OCRv5: A Specialized 5M-Parameter Model Rivaling Billion-Parameter Vision-Language Models on OCR Tasks Paper • 2603.24373 • Published Mar 25
RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild Paper • 2606.23344 • Published Jun 22 • 1
PP-OCRv6: From 1.5M to 34.5M Parameters, Surpassing Billion-Scale VLMs on OCR Tasks Paper • 2606.13108 • Published Jun 11 • 8
Beyond Self-Supervision: A Simple Yet Effective Network Distillation Alternative to Improve Backbones Paper • 2103.05959 • Published Mar 10, 2021 • 1
PP-LCNet: A Lightweight CPU Convolutional Neural Network Paper • 2109.15099 • Published Sep 17, 2021 • 2
PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction Paper • 2503.17213 • Published Mar 21, 2025 • 2
PP-FormulaNet: Bridging Accuracy and Efficiency in Advanced Formula Recognition Paper • 2503.18382 • Published Mar 24, 2025
Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild Paper • 2603.04205 • Published Mar 4 • 2
PP-PicoDet: A Better Real-Time Object Detector on Mobile Devices Paper • 2111.00902 • Published Nov 1, 2021
PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training Paper • 2606.03264 • Published Jun 2 • 27
PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training Paper • 2606.03264 • Published Jun 2 • 27
PaddleOCR-VL-1.5: Towards a Multi-Task 0.9B VLM for Robust In-the-Wild Document Parsing Paper • 2601.21957 • Published Jan 29 • 23
PaddleOCR-VL-1.5: Towards a Multi-Task 0.9B VLM for Robust In-the-Wild Document Parsing Paper • 2601.21957 • Published Jan 29 • 23
PaddleOCR-VL: Boosting Multilingual Document Parsing via a 0.9B Ultra-Compact Vision-Language Model Paper • 2510.14528 • Published Oct 16, 2025 • 129