NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #1619 most downloaded on PyPI
Framework for large language model evaluations
Last release today
22 Sep 2026
Ships on a steady schedule
a new release about every 1 weeks
Nearly every release is documented
notes for 60 of the last 60 stable releases
2 versions withdrawn
withdrawn after publishing
2 years old
246 releases · first in 2024
Fix issue with logs from S3 buckets in inspect view.
sort() method to Dataset (defaults to sorting by sample input length).write_eval_log() now ignores unserializable objects in metadata fields.
write_eval_log() now ignores unserializable objects in metadata fields.read_eval_log() now takes a str or FileInfo (for compatibility w/ list returned from list_eval_logs())..ipynb file due to lack of dependencies (e.g. nbformat).One column per month.
inspect view command for viewing eval log files.
inspect view command for viewing eval log files.Score now has an optional answer field, which denotes the answer text extracted from model output.ValueToFloat function for customising how textual values mapped to float.model_graded_qa more flexible with separate instruction template and grade_pattern, as well providing partial_credit as an option.chain_of_thought() and self_critique() to instruct the model to reply with ANSWER: $ANSWER at the end on its own line.match(numeric=True) (better currency and decimal handling).answer() patterns so that they detect letter and word answers both within and at the end of model output.Plan now has an optional cleanup function which can be used to free per-sample resources (e.g. Docker containers) even in the case of an evaluation error.Dataset.filter method for filtering samples using a predicate.Dataset slices (e.g. dataset[0:100]) now return a Dataset rather than list[Sample].INSPECT_LOG_DIR in .env file is now correctly resolved for execution within subdirectories.inspect list tasks and list_tasks() now only parse source files (rather than loading them), ensuring that it is fast even for task files that have non-trivial global initialisation.inspect list logs and list_eval_logs() now enumerate log files recursively by default, and only enumerate json files that match log file naming conventions.header_only option for read_eval_log() and inspect info log-file for bypassing the potentially expensive reading of samples.filter option for list_eval_logs() to filter based on log file header info (i.e. anything but samples).__main__.py entry point for invocation via python3 -m inspect_ai.ToolDef (renamed to ToolInfo).completion property on ModelOutput with no choices.- Initial release.
Nothing published for this version
Nothing published for this version
Your coding agent can read these notes before it upgrades. Set up the MCP server →