NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #4961 most downloaded on PyPI
A library for performing inference using trained models.
Last release 3 months ago
11 Jun 2026
Ships unpredictably
gaps range from 8 days to 5 months
Nearly every release is documented
notes for 60 of the last 60 stable releases
2 versions withdrawn
withdrawn after publishing
4 years old
124 releases · first in 2022
One column per quarter.
change the default model to yolox, as table output appears to be better and speed is similar to yolox_quantized
chore: remove logger info for chipper since its private
text_as_html attribute and text attribute stores text without html tagsbreaking change: function unstructured_inference.models.tables.recognize no longer takes out_html parameter and it now only returns table cell data fo…
run_prediction callunstructured_inference.models.tables.recognize no longer takes out_html parameter and it now only returns table cell data format (lists of dictionaries)Allow table model to accept optional OCR tokens
Fix: include onnx as base dependency.
Fix a memory leak in DonutProcessor when using large images in numpy format
fix a bug where invalid zoom factor lead to exceptions; now invalid zoom factors results in no scaling of the image
## 0.7.5 * Improved packaging
Dynamic beam search size has been implemented for Chipper, the decoding process starts with a size = 1 and changes to size = 3 if repetitions appear.
Integration of Chipperv2 and additional Chipper functionality, which includes automatic detection of GPU, bounding box prediction and hierarchical rep
Sort elements extracted by pdfminer to get consistent result from aggregate_by_block()
pdfminer to get consistent result from aggregate_by_block()Download yolox already quantized from HF
Remove all OCR related code expect the table OCR code
Stop passing ocr_languages parameter into paddle to avoid invalid paddle language code error, this will be fixed until we have the mapping from standa
Add functionality to keep extracted image elements while merging inferred layout with extracted layout
source property for elements generated by pdfminer.add a function to automatically scale table crop images based on text height so the text height is optimum for tesseract OCR task
tesseract OCR taskconfig.pyfeat: make table transformer parameters configurable by @badGarnet in https://github.com/Unstructured-IO/unstructured-inference/pull/224
Full Changelog: https://github.com/Unstructured-IO/unstructured-inference/compare/0.6.1...0.6.3
feat: add config class by @badGarnet in https://github.com/Unstructured-IO/unstructured-inference/pull/218 This change allows a user to specific infer
yolox the default model for element detection and removes duplicated or near duplicated bounding boxes in the results to reduce noise in the final elements.Full Changelog: https://github.com/Unstructured-IO/unstructured-inference/compare/0.5.31...0.6.1
source property to our elements, so you can know where the information was generated (OCR or detection model)Add functionality to extract and save images from the page
make test fails since paddle doesn't work on M1/M2 chip locallyadd env variable ENTIRE_PAGE_OCR to specify using paddle or tesseract on entire page OCR
ENTIRE_PAGE_OCR to specify using paddle or tesseract on entire page OCRtable structure detection now pads the input image by 25 pixels in all 4 directions to improve its recall
fix a bug in table cell to html conversion where cells spanning multiple rows are not respected in the output
cells_to_html doesn't handle cells spanning multiple rows properlyremove preprocessing for OCR in table structure transformer
Add functionality to bring back embedded images in PDF
Add object-detection classification probabilities to LayoutElement for all currently implemented object detection models
adds safe_division to replace 0 with machine epsilon for float to avoid division by 0
safe_division to replace 0 with machine epsilon for float to avoid division by 0safe_division to area overlap calculations in unstructured_inference/inference/elements.pysafe_division to replae 0 with machine epsilon for float to avoid division by 0safe_division to area overlap calculations in unstructured_inference/inference/elements.py## 0.5.20 * Adds YoloX quantized model
Add functionality to supplement detected layout with elements from the full page OCR
This version adds differentiated thresholds for (optional, alternative model) detectron2_mask_rcnn.
Use OMP_THREAD_LIMIT to improve tesseract performance
OMP_THREAD_LIMIT to improve tesseract performanceFix to no longer create a directory for storing processed images
Handle an uncaught TesseractError
Add TIFF test file and TIFF filetype to test_from_image_file in test_layout
test_from_image_file in test_layoutFix extracted image elements being included in layout merge
Fix a pdfminer error when using process_data_with_model
process_data_with_modelprocess_data_with_modelAdd warning when chipper is used with < 300 DPI
## 0.5.10 * Implement full-page OCR
Handle exceptions from Tesseract
Add alternative architecture for detectron2 (but default is unchanged)
| Library | From | To |
|---|---|---|
| transformers | 4.29.2 | 4.30.2 |
| opencv-python | 4.7.0.72 | 4.8.0.74 |
| ipython | 8.12.2 | 8.14.0 |
hotfix to handle issue storing images in a new dir when the pdf has no file extension
Updated detectron2 version to avoid errors related to deprecated PIL reference
annotate and _get_image_array methods of PageLayout to get the image from the image_path property if the image property is None.image_metadata property to PageLayout & set page.image to None to reduce memory usage.DocumentLayout.from_file to open only one image.load_pdf to return either Image objects or Image paths.Reduced memory usage when working on PDFs
| Library | From | To |
|---|---|---|
| ruff | 0.0.270 | 0.0.276 |
| mypy | 1.3.0 | 1.4.1 |
| onnxruntime | 1.15.0 | 1.15.1 |
pdf2image.convert_from_pathpdf2image.convert_from_pathTweak to element ordering to make it more deterministic
## 0.5.3 * Refactor for large model
Combine inferred elements with extracted elements
Store page numbers when processing PDFs
Preserve image format in PIL.Image.Image when loading
Fixed patches not being a package.
Patch pdfminer.six to fix parsing bug
Output of table extraction is now stored in text_as_html property rather than text property
text_as_html property rather than text propertyAdded the ability to pass ocr_languages to the OCR agent for users who need non-English language packs.
ocr_languages to the OCR agent for users who need
non-English language packs.Added logic to partition granular elements (words, characters) by proximity
Allow extracting tables from higher level functions
Pin protobuf version to avoid errors
Add paddleocr dependency to setup for x86_64 machines
Fixed some cases where image elements were not being OCR'd
Removed control characters from tesseract output
## 0.2.7 * Fixed duplicated load_pdf call
Add YoloX model for images and PDFs
Download default model from huggingface
Your coding agent can read these notes before it upgrades. Set up the MCP server →