RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild Paper • 2606.23344 • Published Jun 22 • 1
PP-OCRv6: From 1.5M to 34.5M Parameters, Surpassing Billion-Scale VLMs on OCR Tasks Paper • 2606.13108 • Published Jun 11 • 8
PP-OCRv5: A Specialized 5M-Parameter Model Rivaling Billion-Parameter Vision-Language Models on OCR Tasks Paper • 2603.24373 • Published Mar 25
PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training Paper • 2606.03264 • Published Jun 2 • 26
PaddleOCR-VL-1.5: Towards a Multi-Task 0.9B VLM for Robust In-the-Wild Document Parsing Paper • 2601.21957 • Published Jan 29 • 23
view article Article PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters PaddlePaddle • Jun 22 • 30
Running Agents Featured 255 PaddleOCR-VL Online Demo 📈 255 Extract text, tables, formulas, and charts from images
Running Agents 82 PP-OCRv5 Online Demo 🌍 82 Universal-Scene Text Recognition Model with High-Accuracy