NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #524 most downloaded on PyPI
Plumb a PDF for detailed information about each char, rectangle, and line.
Last release 3 months ago
15 Jun 2026
Ships fairly regularly
a new release about every 3 months
Nearly every release is documented
notes for 60 of the last 60 stable releases
Nothing withdrawn
no release was ever pulled
11 years old
76 releases · first in 2015
Whoops.
Whoops.
When extracting table cells, use chars' midpoints instead of top-points.
" " charsNothing published for this version
One column per quarter.
Nothing published for this version
Adds Page.extract_words(...), inspired by @jsfenfen's coalesce_words.py
Page.extract_words(...), inspired by @jsfenfen's coalesce_words.pyPage.filter(...)CroppedPage.from_path to .open, and makes PDF class compatible with with statements.atexit)Nothing published for this version
Quickfix to v0.3.0; changes get_text(...) -> extract_text(...) for symmetry's sake.
Quickfix to v0.3.0; changes get_text(...) -> extract_text(...) for symmetry's sake.
A _ton_ of improvements and new features:
A ton of improvements and new features:
pandas requirement and usage.
within_bbox and similar methods, thanks to short-circuiting & operators.floats to Decimals to improve accuracy of equality comparisons.Container, Page, and CroppedPage classes.Page.crop(...).Page.extract_table(...) for Tabula-like functionality.PDF.metadata property.Container.rect_edges and Container.edges, decomposing each rectangle decomposed into its constituent lines.collate_chars(...) to get_text(...) (while retaining a reference to the former).Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Your coding agent can read these notes before it upgrades. Set up the MCP server →