NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #1575 most downloaded on PyPI
SDK and CLI for parsing PDF, DOCX, HTML, and more, to a unified document representation for powering downstream workflows such as gen AI applications.
Last release today
18 Sep 2026
Ships on a steady schedule
a new release about every 9 days
Nearly every release is documented
notes for 60 of the last 60 stable releases
Nothing withdrawn
no release was ever pulled
2 years old
217 releases · first in 2024
One column per month.
Export to WebVTT format (#3036) (`d276e60`)
859c302)6198e69)477359b)90ce93d)a3d2b4b)Create a backend parser for XBRL instance reports (#3017) (`334ba6e`)
Security vulnerabilities with XML External Entity and related attacks (#3009) (`576bada`)
Inference engines abstraction for object detection model family with HF Transformers and ONNX runtime (#2959) (`14e474c`)
14e474c)e6ccb8b)d4c8713)9721321)ae4fdbb)Add chart extraction models (#2848) (`fe45c71`)
Webvtt and source tracker (#2787) (`0602a7c`)
Drop support for Python 3.9 (#2905) (`7f38658`)
Off-by-one error for page indexing in vlm_pipeline (#2902) (`08f49e2`)
New picture classifier v2.0 (#2889) (`43badc3`)
43badc3)ac16a26)00273f6)1b4d82d)2fe9def)daf2bc6)Support for DeepSeek-OCR in VLM pipeline (#2798) (`19af03f`)
Enrichment annotations in the new meta format (#2859) (`aab3ff5`)
aab3ff5)2b83fdd)cbc6537)d9295df)a0530a2)3ef4525)ffafe58)ed57089)Add preset for using granite-docling via vllm and other apis (#2792) (`241d19e`)
Add YAML output format to CLI (#2768) (`da7678a`)
Clear word/char cells when force_full_page_ocr is used (#2738) (`1df0560`)
1df0560)edbabfc)609069d)d007ba0)aebe25c)c97715f)experimental: Add experimental TableCropsLayoutModel (#2669) (`1344362`)
1344362)ad97e52)6ef4ffd)54cd6d7)e580554)examples: Remove deprecation warnings with export_to_dataframe (#2638) (`f552862`)
2087c6b)6fb9a5f)463a3fd)da4c2e9)ce5a099)b216ad8)03e7c7d)8af228f)d549445)ac9fc58)f552862)Add the Image backend (#2627) (`3495b73`)
Default to EasyOCR in Python 3.14 (#2605) (`5c27567`)
vlm: Track generated tokens and stop reasons for VLM models (#2543) (`6a04e27`)
Extract response from api_image_request in picture description (#2571) (`8360aa5`)
Use threading in the standard pipeline and move old behavior to legacy (#2452) (`268d027`)
vlm: Add num_tokens as attribtue for VlmPrediction (#2489) (`b6c892b`)
pdf: Support for password-protected PDF documents (#2499) (`bbe82a6`)
bbe82a6)a30e6a7)657ce8b)4227fcc)b66624b)docx: Process drawingml objects in docx (#2453) (`1682993`)
Avoid downloading easyocr models by default (#2454) (`688a7df`)
AutoOCR model selecting the best OCR model available and deprecating the usage of EasyOCR (#2391) (`f7244a4`)
f7244a4)f11f8c0)db985bb)cce18b2)ee55013)b5f7fef)0610d01)9705f40)markdown: Setext heading support (#2359) (`ee73ffa`)
Repetition-based StoppingCriteria for GraniteDocling (#2323) (`1e9dc43`)
1e9dc43)c803abe)68ae7cc)654c70f)9d67bb9)Rich tables for MSWord backend (#2291) (`e2482a2`)
Add granite-docling model (#2272) (`17afb66`)
Address deprecation warnings of dependencies (#2237) (`c696549`)
Updating default parameters to get better performance with docling-parse (#2208) (`b49d1ad`)
Heron layout model as new default (#1971) (`e38aa0f`)
[Beta] Extraction with schema (#2138) (`9f4bc5b`)
9f4bc5b)a283ccf)4d94e38)9f0286b)9904d14)Upgrade to RapidOCR 3.x (#2088) (`3f60a0f`)
Vllm extra only for linux x86_64 (#2126) (`488f6cd`)
CLI: Option to download arbitrary HuggingFace model (#2123) (`cdf079d`)
New code formula model (#2042) (`d2494da`)
Add backend for METS with Google Books profile (#1989) (`31087f3`)
Add convert_string to document-converter (#2069) (`b09033c`)
HTML: Concatenation of child strings in table cells and list items (#1981) (`5132f06`)
Add option to control empty clusters in layout postprocessing (#1940) (`a436be7`)
cca05c4)e1e3053)95e7096)c5fb353)Layout model specification and multiple choices (#1910) (`2b8616d`)
Introduce LayoutOptions to control layout postprocessing behaviour (#1870) (`ec6cf6f`)
ec6cf6f)56a0e10)598c9c5)ae39a94)Leverage new list modeling, capture default markers (#1856) (`0533da1`)
Updated granite vision model version for picture description (#1852) (`d337825`)
Support audio input (#1763) (`1557e7c`)
1557e7c)861abcd)215b540)d26dac6)1350a8d)dd7f64f)dbab30e)Make Page.parsed_page the only source of truth for text cells, add OCR cells to it (#1745) (`7d3302c`)
7d3302c)df14022)f28d23c)b886e4d)7a275c7)6613b9e)e979750)f7f3113)9dbcb3d)a2b83fe)Remove typer and click constraints (#1707) (`8846f1a`)
Simplify dependencies, switch to uv (#1700) (`cdd4018`)
Add visualization of bbox on page with html export. (#1663) (`b356b33`)
ocr: Auto-detect rotated pages in Tesseract (#1167) (`45265bf`)
Add textbox content extraction in msword_backend (#1538) (`12a0e64`)
f4d9d41)0e00a26)f2e9c07)98b5eeb)Improve parallelization for remote services API calls (#1548) (`3a04f2a`)
AsciiDoc header identification (#1562) (#1563) (`4046d0b`)
Add smoldocling in download utils (#1577) (`127e386`)
127e386)776e7ec)f1658ed)7c70573)4ab7e9d)cc45396)976e92e)94d66a0)Your coding agent can read these notes before it upgrades. Set up the MCP server →