NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #3519 most downloaded on PyPI
🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & scraper
Last release 11 days ago
23 Sep 2026
Ships fairly regularly
a new release about every 3 weeks
Most releases are documented
notes for 37 of the last 60 stable releases
Nothing withdrawn
no release was ever pulled
2 years old
72 releases · first in 2024
One column per month.
Added comprehensive system diagnostics tool
New Doctor Feature
Dockerized API Server
Managed Browser Integration
ManagedBrowser class for better browser lifecycle managementEnhanced HTML Processing
Browser Handling
Database Management
Resource Management
When upgrading to v0.3.73, be aware of the following changes:
Docker Deployment:
If using custom browser management:
For database operations:
Using the Doctor:
crawl4ai doctorNew ContentCleaningStrategy class:
ContentCleaningStrategy class:
proxy_config option for authenticated proxy connectionsfit_markdown: Optimized markdown output with main content focusfit_html: Clean HTML with only essential contentsrc, data-src, srcset, etc.)ContentCleaningStrategy uses configurable thresholds for customizationOverlappingWindowChunking: Allows for overlapping chunks of text, useful for maintaining context between chunks.
OverlappingWindowChunking: Allows for overlapping chunks of text, useful for maintaining context between chunks.SlidingWindowChunking: Improved to handle edge cases and last chunks more effectively.CHUNK_TOKEN_THRESHOLD in config to 2048 tokens (2^11) for better compatibility with most LLM models.AsyncPlaywrightCrawlerStrategy.close() method to use a shorter sleep time (0.5 seconds instead of 500), significantly reducing wait time when closing the crawler.CosineStrategy:
load_HF_embedding_model function, allowing for easier swapping of embedding models.JsonCssExtractionStrategy and JsonXPathExtractionStrategy for better JSON-based extraction.CosineStrategy, now using the automatically detected device.quickstart_async.py for generating a knowledge graph from crawled content.These updates aim to provide more flexibility in text processing, improve performance, and enhance the overall capabilities of the crawl4ai library. The new chunking strategies, in particular, offer more options for handling large texts in various scenarios.
Nothing published for this version
Implemented playwright_stealth for improved bot detection avoidance.
Enhanced Browser Stealth:
playwright_stealth for improved bot detection avoidance.StealthConfig for fine-tuned control over stealth parameters.User Simulation:
simulate_user option to mimic human-like interactions (mouse movements, clicks, keyboard presses).Navigator Override:
override_navigator option to modify navigator properties, further improving bot detection evasion.Improved iframe Handling:
process_iframes parameter to extract and integrate iframe content into the main page.Flexible Browser Selection:
Include Links in Markdown:
include_links_on_markdown in crawl method.Better Error Handling:
Image Processing Enhancements:
Crawling Flexibility:
delay_before_return_html parameter.Performance Optimization:
crawl_with_user_simulation() demonstrating the use of user simulation and navigator override features.New Hook: Added before_retrieve_html hook in AsyncPlaywrightCrawlerStrategy.
before_retrieve_html hook in AsyncPlaywrightCrawlerStrategy.delay_before_return_html parameter to allow waiting before retrieving HTML content.
smart_wait function now uses page_timeout (default 60 seconds) instead of a fixed 30-second timeout.
page_timeout=your_desired_timeout (in milliseconds) when calling crawler.arun().browser_type="firefox" or browser_type="webkit" when initializing AsyncWebCrawler.screenshot=True when calling crawler.arun().extra_args parameter.LLMExtractionStrategy.process_iframes=True in the crawl method.get_delayed_content method in AsyncCrawlResponse.result.get_delayed_content(delay_in_seconds) after crawling.perform_completion_with_backoff function now supports additional arguments for more customized API calls to LLM providers.quickstart_async.py with examples of:
These updates significantly enhance the flexibility, accuracy, and robustness of crawl4ai, providing users with more control and options for their web crawling and content extraction tasks.
Enhance AsyncWebCrawler with smart waiting and screenshot capabilities
Enhance AsyncWebCrawler with smart waiting and screenshot capabilities
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Your coding agent can read these notes before it upgrades. Set up the MCP server →