NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #1751 most downloaded on PyPI
SDK for integrating Braintrust
Last release today
17 Sep 2026
Ships on a steady schedule
a new release about every 2 weeks
Some releases are documented
notes for 17 of the last 60 stable releases
1 version withdrawn
withdrawn after publishing
3 years old
284 releases · first in 2023
One column per quarter.
Added DSPy integration with wrap_dspy wrapper for automatic tracing of DSPy applications
Added DSPy integration with wrap_dspy wrapper for automatic tracing of DSPy applications
Added OpenTelemetry distributed tracing helpers ( context_from_span_export() and span_context_from_span_export() ) for seamless trace propagation between Braintrust and OpenTelemetry across service boundaries
Added support for GEMINI_API_KEY environment variable
Ensure experiments use SpanComponentsV3 by default
Added OpenTelemetry compatibility mode for seamless integration between Braintrust and OTEL tracing
Added OpenTelemetry compatibility mode for seamless integration between Braintrust and OTEL tracing
Added setup_claude_agent_sdk for automatic tracing of Claude Agent SDK applications
Improved Anthropic wrapper to log consistent input/output format
Added strict parameter to Prompt.build for strict schema validation
Added SpanComponentsV4 support
Nothing published for this version
Nothing published for this version
Nothing published for this version
Python SDK now correctly nests spans logged from inside tool calls in OpenAI Agents
Support data masking (see docs )
Support data masking (see docs )
Remote evals in Python SDK
Support tags in Eval hooks
Validate attachment file readability at creation time
Allow non-batch span processors in BraintrustSpanProcessor
Fix openai-agents to inherit the right tracing context
Added environment parameter to load_prompt
Added environment parameter to load_prompt
The Otel SpanProcessor now keeps traceloop.* spans by default
Experiments can now be run without sending results to the server
Span creation is significantly faster in Python
Fix langchain-py integration tracing when users use a @traced method
Fix langchain-py integration tracing when users use a @traced method
Wrap OpenAI responses.parse
Add @traced support for generator functions
Autoevals PY (version 0.0.130)
When running multiple trials per input ( trial_count > 1 ), you can now access the current trial index (0-based) via hooks.trialIndex in your task fun
When running multiple trials per input ( trial_count > 1 ), you can now access the current trial index (0-based) via hooks.trialIndex in your task function
Added BraintrustExporter in addition to BraintrustSpanProcessor
Bound max ancestors in git to 1,000
Added BraintrustSpanProcessor to simplify Braintrust’s integration with OpenTelemetry
Added support for loading prompts by ID via the load_prompt function. You can now load prompts directly by their unique identifier
Nothing published for this version
The SDK’s under-the-hood log queue will not block when full and has a default size of 25000 logs
The SDK’s under-the-hood log queue will not block when full and has a default size of 25000 logs
You can configure the max size by setting BRAINTRUST_LOG_QUEUE_MAX_SIZE in your environment
Improvements to the logging of parallel tool calls
Attachments are now converted to base64 data URLs, making it easier to work with image attachments in prompts
Add project.publish() to directly push prompts to Braintrust (without running braintrust push )
Add project.publish() to directly push prompts to Braintrust (without running braintrust push )
@traced now works correctly with async generator functions
The OpenAI and Anthropic wrappers set provider metadata
Improve retry logic in the control plane connection (used to create new experiments and datasets)
Added support for metadata and tags arguments to invoke
Added support for metadata and tags arguments to invoke
The SDK now gracefully handles OpenAI’s NotGiven parameter
Added span.link() to synchronously generate permalinks
Added BraintrustSpanProcessor to simplify integration with OpenTelemetry
Fix a bug where large experiments would drop spans if they could not flush data fast enough
Fix a bug where large experiments would drop spans if they could not flush data fast enough
Fix bug in attachment uploading in evals executed with npx braintrust eval
Upgrading zod dependency from ^3.22.4 to ^3.25.3
Added support for loading prompts by ID via the loadPrompt function
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Your coding agent can read these notes before it upgrades. Set up the MCP server →