NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #3575 most downloaded on PyPI
Awesome multilingual OCR and document parsing toolkits based on PaddlePaddle
Last release 3 months ago
11 Jun 2026
Release timing varies
gaps range from 9 days to 2 months
Some releases are documented
notes for 23 of the last 60 stable releases
2 versions withdrawn
withdrawn after publishing
6 years old
65 releases · first in 2020
One column per quarter.
Accuracy boost : Medium tier achieves +4.6% detection and +5.1% recognition over PP-OCRv5_server, surpassing mainstream VLMs (Qwen3-VL-235B, GPT-5.5)
Release PP-OCRv6
发布 PP-OCRv6
Full Changelog: v3.6.0...v3.7.0
Release the PaddleOCR-VL-1.6 document parsing solution.
Release the PaddleOCR-VL-1.6 document parsing solution.
Release official PaddleOCR API SDKs for Python, Go, and TypeScript, enabling more convenient integration with PaddleOCR official asynchronous APIs.
Add support for parsing multi-page TIFF files.
发布 PaddleOCR-VL-1.6 文档解析方案。
发布 PaddleOCR 官方 API Python、Go、TypeScript SDK,支持更加便捷地调用 PaddleOCR 官方异步 API。
新增对多页 TIFF 文件解析的支持。
Full Changelog: v3.5.0...v3.6.0
Deeply integrated with the Hugging Face ecosystem, with 20 major models supporting Transformers as the inference backend. Supports flexible switching
Full Changelog: v3.4.1...v3.5.0
PaddleOCR-VL adds llama-cpp-server backend support.
llama-cpp-server backend support.llama-cpp-server 后端支持。Full Changelog: v3.4.0...v3.4.1
Release the PaddleOCR-VL-1.5 complex document parsing solution.
Release the PaddleOCR-VL-1.5 complex document parsing solution.
PaddleOCR-VL-1.5 is a new iterative version of the PaddleOCR-VL series. Based on comprehensive optimization of the core capabilities of version 1.0, the model achieves 94.5% accuracy on the authoritative document parsing benchmark OmniDocBench v1.5, surpassing top global general-purpose large models and document parsing–specific models.
PaddleOCR-VL-1.5 innovatively supports irregular-shaped bounding box localization of document elements, enabling excellent performance in real-world application scenarios such as scanning, skew, warping, screen-photography, and complex illumination, achieving comprehensive SOTA performance. In addition, the model further integrates seal recognition and spotting tasks, with key metrics continuing to lead mainstream models.
You can use it online on the PaddleOCR official website or call the model API.
Add support for calling MLX-VLM inference services.
PaddleOCR-VL now supports cross-page table merging and multi-level heading reconstruction.
PP-StructureV3 adds support for the format_block_content and markdown_ignore_labels parameters.
Fixed an issue where accessing the /docs endpoint in the official PaddleOCR-VL image would result in an error.
发布 PaddleOCR-VL-1.5 复杂文档解析方案。
PaddleOCR-VL-1.5 是 PaddleOCR-VL 系列的全新迭代版本。在全面优化 1.0 版本核心能力的基础上,该模型在文档解析权威评测集 OmniDocBench v1.5 上斩获了 94.5% 的高精度,超越了全球的顶尖通用大模型及文档解析专用模型。
PaddleOCR-VL-1.5 创新性地支持了文档元素的异形框定位,使得 PaddleOCR-VL-1.5 在扫描、倾斜、弯折、屏幕拍摄及复杂光照等真实落地场景中均表现卓越,实现了全面的 SOTA。此外,模型进一步集成了印章识别与文本检测识别任务,关键指标持续领跑主流模型。
您可以在 PaddleOCR官网 在线使用或者调用该模型的API。
新增对 MLX-VLM 推理服务的调用支持。
PaddleOCR-VL 支持合并跨页表格、多级标题重建功能。
PP-StructureV3 支持 format_block_content、markdown_ignore_labels 参数。
修复 PaddleOCR-VL 官方镜像访问 /docs 接口报错的问题。
Full Changelog: v3.3.3...v3.4.0
PaddleOCR-VL now supports specifying custom model names and API keys, and can seamlessly integrate with inference services from third-party platforms
Full Changelog: v3.3.2...v3.3.3
2025.11.13 v3.3.2 released Full Changelog : v3.3.1...v3.3.2
Full Changelog: v3.3.1...v3.3.2
Fixed the issue where the document image preprocessing switch did not take effect in PP-StructureV3 and PaddleOCR-VL.
Full Changelog: v3.3.0...v3.3.1
PaddleOCR-VL is a SOTA and resource-efficient model tailored for document parsing. Its core component is PaddleOCR-VL-0.9B, a compact yet powerful vis
Released PaddleOCR-VL:
Model Introduction:
Core Features:
Released PP-OCRv5 Multilingual Recognition Model:
发布PaddleOCR-VL:
模型介绍:
特性:
发布PP-OCRv5小语种识别模型:
Introduced training, inference, and deployment for PP-OCRv5 recognition models in English, Thai, and Greek. The PP-OCRv5 English model delivers an 11%
Significant Model Additions:
Deployment Capability Upgrades:
Benchmark Support:
Bug Fixes:
use_chart_parsing) in the PP-StructureV3 configuration files compared to other pipelines.Other Enhancements:
重要模型新增:
部署能力升级:
Benchmark支持:
Bug修复:
use_chart_parsing 等开关行为与其他产线不统一的问题。其他升级:
Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v3.1.1...v3.2.0
Added the missing methods save_vector, save_visual_info_list, load_vector, and load_visual_info_list in the PP-ChatOCRv4 class.
Bug Fixes:
save_vector, save_visual_info_list, load_vector, and load_visual_info_list in the PP-ChatOCRv4 class.glossary and llm_request_interval to the translate method in the PPDocTranslation class.Documentation Improvements:
Others:
puremagic instead of python-magic to reduce installation issues.bug修复:
PP-ChatOCRv4 类缺失的save_vector、save_visual_info_list、load_vector、load_visual_info_list 方法。PPDocTranslation 类的 translate 方法缺失的 glossary 和 llm_request_interval 参数。文档优化:
其他:
puremagic 代替 python-magic,减少安装问题。Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v3.1.0...v3.1.1
Added PP-OCRv5 Multilingual Text Recognition Model, which supports the training and inference process for text recognition models in 37 languages, inc
Key Models and Pipelines:
New MCP server: Details
Documentation Optimization: Improved the descriptions in some user guides for a smoother reading experience.
重要模型和产线:
新增MCP server:详情
文档优化: 优化了部分使用文档描述,提升阅读体验。
修复enable_mkldnn参数不生效的问题,恢复CPU默认使用MKL-DNN推理的行为。
enable_mkldnn参数不生效的问题,恢复CPU默认使用MKL-DNN推理的行为。模型默认下载源从BOS改为HuggingFace,同时也支持用户通过更改环境变量PADDLE_PDX_MODEL_SOURCE为BOS,将模型下载源设置为百度云对象存储BOS。
功能新增:
BOS改为HuggingFace,同时也支持用户通过更改环境变量PADDLE_PDX_MODEL_SOURCE为BOS,将模型下载源设置为百度云对象存储BOS。Bug修复:
export_paddlex_config_to_yaml无法正常工作的问题。文档优化:
enable_mkldnn参数的说明,使其更准确地描述程序的实际行为。lang和ocr_version参数描述的错误。其他:
更新 PP-OCRv5默认模型配置,检测和识别均由mobile改为server模型。为了改善大多数的场景默认效果,配置中的参数limit_side_len由736改为64
limit_side_len由736改为64PP-LCNet_x1_0_textline_ori模型,精度99.42%,OCR、PP-StructureV3、PP-ChatOCRv4产线的默认文本行方向分类器改为该模型PP-LCNet_x0_25_textline_ori模型,精度提升3.3个百分点,当前精度98.85%use_textline_orientation参数。FatalError: Process abort signal is detected by the operating system错误的问题PPStructureV3.concatenate_markdown_pages方法不存在的问题。paddleocr.PaddleOCR时同时指定lang和model_name时model_name不生效的问题。发布全场景文字识别模型PP-OCRv5: 单模型支持五种文字类型和复杂手写体识别;整体识别精度相比上一代提升13个百分点。
发布全场景文字识别模型PP-OCRv5: 单模型支持五种文字类型和复杂手写体识别;整体识别精度相比上一代提升13个百分点。
发布通用文档解析方案PP-StructureV3: 支持多场景、多版式 PDF 高精度解析,在公开评测集中领先众多开源和闭源方案。
发布智能文档理解方案PP-ChatOCRv4: 原生支持文心大模型4.5 Turbo,精度相比上一代提升15个百分点。
重构部署能力,统一推理接口: PaddleOCR 3.0 融合了飞桨 PaddleX3.0 工具的底层能力,全面升级推理、部署模块,优化 2.x 版本的设计,统一并优化了 Python API 和命令行接口(CLI)。部署能力现覆盖高性能推理、服务化部署及端侧部署三大场景。
适配飞桨框架 3.0,优化训练流程: 新版本已兼容飞桨 3.0 的 CINN 编译器等最新特性,静态图模型存储文件名由 xxx.pdmodel 改为 xxx.json。
统一模型名称: 对PaddleOCR3.0支持的模型命名体系进行了更新,采用更规范、统一的命名规则,为后续迭代与维护奠定基础。
update docs by @cuicheng01 in https://github.com/PaddlePaddle/PaddleOCR/pull/14031
CMAKE_CXX_FLAGS optimize flag by @Hirozy in https://github.com/PaddlePaddle/PaddleOCR/pull/14059create_predictor function to accept array of ONNX Execution Providers by @Salmondx in https://github.com/PaddlePaddle/PaddleOCR/pull/14078recovery_to_markdown.py by @Coobiw in https://github.com/PaddlePaddle/PaddleOCR/pull/14216rec_image_shape when manually set by @JesuisTong in https://github.com/PaddlePaddle/PaddleOCR/pull/14371OverflowError in text_visual func by @GreatV in https://github.com/PaddlePaddle/PaddleOCR/pull/14758paddleocr --image_dir xxx.png --lang ch_doc
Similarly, in the Python API, you only need to change the lang parameter to ch_doc.
Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v2.9.1...v2.10.0
[cherry-pick] update paddle2onnx doc by @inisis in https://github.com/PaddlePaddle/PaddleOCR/pull/14051
Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v2.9.0...v2.9.1
fix: table recognition content is not escaped properly by @GreatV in https://github.com/PaddlePaddle/PaddleOCR/pull/13277
Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v2.8.0...v2.9.0
[cherry-pick] add project url and fix a bug by @GreatV in https://github.com/PaddlePaddle/PaddleOCR/pull/13281
Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v2.8.0...v2.8.1
[Cherry-pick] #10515 by @ToddBear in https://github.com/PaddlePaddle/PaddleOCR/pull/10537
cls_x and bbox_x is possibly unbound by @SigureMo in https://github.com/PaddlePaddle/PaddleOCR/pull/10991try_import from paddle.utils by @neteroster in https://github.com/PaddlePaddle/PaddleOCR/pull/11820np.int by @Liyulingyue in https://github.com/PaddlePaddle/PaddleOCR/pull/12249slice op demo for quickstart by @GreatV in https://github.com/PaddlePaddle/PaddleOCR/pull/12439Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v2.7.5...v2.8.0
Nothing published for this version
This release contains the missed commits from v2.7.0 to v2.7.1. fixed : #11824
This release contains the missed commits from v2.7.0 to v2.7.1. fixed : #11824
## What's Changed fixed #11808
fixed #11808
add finnish language files by @savikko in https://github.com/PaddlePaddle/PaddleOCR/pull/10850
cls_x and bbox_x is possibly unbound by @SigureMo in https://github.com/PaddlePaddle/PaddleOCR/pull/10973Full Changelog: https://github.com/PaddlePaddle/PaddleOCR/compare/v2.7.0...v2.7.2
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Your coding agent can read these notes before it upgrades. Set up the MCP server →