NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #2982 most downloaded on PyPI
Open reproduction of consastive language-image pretraining (CLIP) and related.
Last release 7 months ago
27 Feb 2026
Ships fairly regularly
a new release about every 2 months
Most releases are documented
notes for 40 of 55 stable releases
Nothing withdrawn
no release was ever pulled
5 years old
55 releases · first in 2022
One column per quarter.
Nothing published for this version
Remove non-existent MetaCLIP 2 L/14 checkpoint by @voidism in #1111
Full Changelog: v3.1.0...v3.2.0
Add support for MetaCLIP2 WorldWide models by @rwightman in #1100
Full Changelog: v3.0.0...v3.1.0
Initial work on adding local-dir: schema for model & tokenizer loading from local folder by @rwightman in #1069
Full Changelog: v2.32.0...v3.0.0
API for getting intermediate image and text features, forward_intermediates() by @rwightman in #1035
Full Changelog: v2.31.0...v2.32.0
Add SigLIP2 models by @rwightman in #1033
Support using timm optimizers for alternative to adamw default by @rwightman in #979
Full Changelog: v2.29.0...v2.30.0
All default pretrained weights pushed to HF hub by @rwightman in #970
Full Changelog: v2.28.0...v2.29.0
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Fix missing space in error message
Stop checking model name when loading CoCa models
Nothing published for this version
* Add EVA models * Support serial worker training * Fix Python 3.7 compatibility
* Add DataComp models
Enable int8 inference without .weight attribute
.weight attribute* Update push_to_hf_hub
Nothing published for this version
Fixes for context_length and vocab_size attributes
Fixes for context_length and vocab_size attributes
* Add improved g-14 weights * Update protobuf version
Add samples per second per gpu logging
Move dataset mixtures logic to shard level
Add support for dataset mixtures with different sampling weights
Updated convnext configs for consistency
Add MSCOCO CoCa finetunes to pretrained models
* coca support and weights * ConvNeXt-Large weights
hf-hub:org/model_id support for loading models w/ config and weights in Hugging Face Hub
hf-hub:org/model_id support for loading models w/ config and weights in Hugging Face HubAdded an up-to-date example slurm script for large training jobs.
base & base_w pretrained models addedtimm- model prefix removed from configstimm augmentation + regularization (dropout / drop-path) supportedFix wandb collapsing multiple parallel runs into a single one
Fix braceexpand memory explosion for complex webdataset urls
* Fix release
wrapped patchdropout in a torch.nn.Module
override the default patch dropout value in 'vision_cfg'
add support for gradient accumulation
add multilingual H/14 xlm roberta large
* fix setup.py _read_reqs
pretrained B/32 xlm roberta base: first multilingual clip trained on laion5B
Add missing hf_tokenizer_name in CLIPTextCfg.
Fix #211, missing RN50x64 config. Fix type of dropout param for ResNet models
Implement grad checkpointing for hf model.
Generalizable Text Transformer with HuggingFace Models (@iejMac)
Add checksum verification for pretrained model weights
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
Nothing published for this version
add missing openai RN50x64 model
* ViT-B/16+ * Add grad checkpointing support * more robust data loader
Nothing published for this version
Nothing published for this version
Nothing published for this version
Your coding agent can read these notes before it upgrades. Set up the MCP server →