PP-OCRv6 is the latest PaddleOCR model family for real-world text detection and recognition, scaling from 1.5M to 34.5M parameters across tiny, small, and medium tiers. The medium and small models support 50 languages including Chinese, Japanese, and 46 Latin-script languages. Key architectural improvements include RepLKFPN for multi-scale text detection and EncoderWithLightSVTR for recognition. Compared to PP-OCRv5_server, the medium tier improves detection Hmean by +4.6 points and recognition accuracy by +5.1 points. Models are available on Hugging Face Hub in safetensors, Paddle inference, and ONNX formats, with support for Transformers, ONNX Runtime, and Paddle Inference backends.
Table of contents
What’s new in PP-OCRv6Quick start with PaddleOCRAvailable inference backendsConclusion710 Impressions