NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #1980 most downloaded on PyPI
ModelScope: bring the notion of Model-as-a-Service to life.
Last release 2 days ago
15 Sep 2026
Ships fairly regularly
a new release about every 3 weeks
Nearly every release is documented
notes for 54 of the last 60 stable releases
Nothing withdrawn
no release was ever pulled
4 years old
93 releases · first in 2022
One column per quarter.
add rife-video-frame-interpolation and model
| 序号 | 模型名称&快捷链接 |
|---|---|
| 1 | 支持qwen1.5系列模型 |
| 2 | RIFE视频插帧 |
| 3 | VFI-RAFT视频插帧 |
| 4 | 轻量级快速图像特征点匹配 |
Nothing published for this version
| 4 | TinyLlama-1.1B-Chat-v1.0 |
| 11 | LanguageBind_Video_merge |
Nothing published for this version
LLMPipeline 支持Swift adapter模型推理
freeU method.work_dir setting in trainer was not taking effect. #573Support int4 model for llm_pipeline
增加3dhuman render and animation 模型
解决新版本transformers position_ids兼容性问题
| 5 | WizardCoder-Python-13B-V1.0 |
Nothing published for this version
Nothing published for this version
check flash attention installation even if use_fast_att is set True
bugfix for qwen
| 4 | 语音合成-越南语-通用领域-24k-发音人tien |
Nothing published for this version
增加baichuan和chatglm2 lora agent示例
| 1 | 读光-文字识别-轻量化端侧识别模型-中英-通用领域 |
| 序号 | 模型名称&快捷链接 |
|---|---|
| 1 | 读光-文字识别-轻量化端侧识别模型-中英-通用领域 |
| 2 | 读光-文字检测-轻量化端侧DBNet行检测模型-中英-通用领域 |
| 3 | CAM++说话人转换点定位-两人-中文 |
| No | Model Name & Link |
|---|---|
| 1 | cv_LightweightEdge_ocr-recognitoin-general_damo |
| 2 | cv_proxylessnas_ocr-detection-db-line-level_damo |
| 3 | speech_campplus-transformer_scl_zh-cn_16k-common |
支持Flextrain training args和push_to_hub
| 达摩院 | ERes2Net说话人确认-英文-VoxCeleb-16k-离线-pytorch | 否 |
该版本共新增上架5个模型。
| 贡献组织 | 模型名称 | 是否支持Finetune |
|---|---|---|
| 达摩院 | ERes2Net说话人确认-英文-VoxCeleb-16k-离线-pytorch | 否 |
| 达摩院 | mPLUG-Owl-多模态对话-英文-7B | 否 |
| 达摩院 | FastInst快速实例分割 | 否 |
| 达摩院 | TransFace人脸识别模型 | 否 |
| 达摩院 | Regularized DINO说话人确认-英文-VoxCeleb-16k-离线-pytorch | 否 |
Nothing published for this version
Nothing published for this version
| 3 | CAM++说话人确认-英文-VoxCeleb-16k |
| 序号 | 模型名称&快捷链接 |
|---|---|
| 1 | ResNet50行人结构化属性识别模型 |
| 2 | DamoFD人脸检测关键点模型-0.5G |
| 3 | CAM++说话人确认-英文-VoxCeleb-16k |
| 4 | 一种具有自我评估能力的机器翻译-中英-通用领域-large |
| No | Model Name & Link |
|---|---|
| 1 | ResNet50 pedestrian-attribute-recognition image |
| 2 | DamoFD face-detection 0.5G |
| 3 | Speech cam++ English-VoxCeleb-16k |
| 4 | Canmt translation with self evaluation zh2en-large |
Nothing published for this version
| 序号 | 模型名称&快捷链接 | 贡献组织 | 是否支持finetune |
| 序号 | 模型名称&快捷链接 | 贡献组织 | 是否支持finetune |
|---|---|---|---|
| 1 | ChatGLM-中英对话大模型-6B | 智谱.AI | |
| 2 | GLM130B-中英大模型 | 智谱.AI | |
| 3 | unidiffuser-v1 | 清华TSAIL | |
| 4 | 元语功能型对话大模型v2 | 元语智能 | |
| 5 | 盘古α 2.6B | 鹏城实验室 | |
| 6 | openjourney | 个人开发者-dienstag | |
| 7 | Rwkv-4-pile-14b | 个人开发者-Blink_DL | |
| 8 | SiameseUIE通用信息抽取-中文-base | ||
| 9 | SiameseUniNLU零样本通用自然语言理解-中文-base · 模型库 (modelscope.cn) |
| No | Model Name & Link | Org | Finetune supported |
|---|---|---|---|
| 1 | ChatGLM-English&Chinese-6B | ZhiPu.AI | |
| 2 | GLM130B-LLM English&Chinese | ZhiPu.AI | |
| 3 | unidiffuser-v1 | TsingHua TSAIL | |
| 4 | ChatYuan-large-v2 | YuanYu | |
| 5 | OpenICommunity/pangu_2_6B | PengCheng Lab | |
| 6 | openjourney | personal-dienstag | |
| 7 | Rwkv-4-pile-14b | personal-Blink_DL | |
| 8 | SiameseUIE information extraction-Chinese-base | ||
| 9 | SiameseUniNLU zero-shot NLU Chinese base model |
-Support run text generation pipeline with args
Nothing published for this version
该小版本共新增上架6个模型,其中新增2个模型支持finetune能力。
该小版本共新增上架6个模型,其中新增2个模型支持finetune能力。
| 序号 | 模型名称&链接 | 支持finetune |
|---|---|---|
| 1 | ControlNet可控图像生成 | |
| 2 | 兰丁宫颈细胞AI辅助诊断模型 | |
| 3 | 读光-文字检测-DB行检测模型-中英-通用领域 | |
| 4 | SOND说话人日志-中文-alimeeting-16k-离线-pytorch | |
| 5 | NeRF快速三维重建模型 | √ |
| 6 | DCT-Net人像卡通化 | √ |
This minor version adds a total of six new models, including two models with finetuning capability.
Nothing published for this version
该版本共新增上架51个模型,其中11个模型支持finetune能力。
该版本共新增上架51个模型,其中11个模型支持finetune能力。
提供finetune的示例脚本,允许用户通过运行脚本命令行传参方式进行模型训练,详细可以参考github脚本
NLP领域新增了backbone + head的开发支持,允许用户任意组合已有的backbone(Encoder) 和任务head,方便在特定任务上切换不同模型进行建模,详细参考文档
贡献者文档完善模型贡献部分,详细参考接入流程概览
数据集接口支持本地文件直接加载 MsDataset.load('/to/path/abc.csv')
模型导出支持nlp_structbert_zero-shot、 nlp_csanmt_translation系列模型
更多SDK功能和变更可查看:https://github.com/modelscope/modelscope/releases/tag/v1.3.0
最后,我们还推出许多任务级别和模型级别的最佳实践教程文档,旨在帮助开发者更好地理解和应用模型。
欢迎关注我们的开源社区:https://github.com/modelscope/modelscope
import torchaudio to avoid unnecessary requirements in frameworkfunasr版本升级 & 语音识别、说话人确认、标点预测增加额外参数配置
import torchaudio to avoid unnecessary requirements in framework该版本共新增上架38个模型,其中14个模型支持finetune能力。
该版本共新增上架38个模型,其中14个模型支持finetune能力。
高性能检测热门应用系列, 基于精度和速度均超越当前经典YOLO系列、面向工业落地的高性能检测框架DAMOYOLO,新增实时口罩检测模型、实时安全帽检测模型、实时人体检测模型、实时香烟检测模型上线,提供开箱即用的高效体验
语音识别、语音合成以及语音唤醒可以基于Modelscope Python SDK进行模型finetune
语音合成,新增方言模型四川话、广东粤语与上海话,新增俄语与韩语外语模型
SambertHifigan语音合成-四川话-通用领域-16k-发音人chuangirl, 方言四川话女声模型
SambertHifigan语音合成-广东粤语-通用领域-16k-发音人jiajia, 方言广东话女声模型
SambertHifigan语音合成-上海话-通用领域-16k-发音人xiaoda, 方言上海话女声模型
SambertHifigan语音合成-俄语-通用领域-16k-发音人masha, 俄语女声模型
SambertHifigan语音合成-韩语-通用领域-16k-发音人kyong, 韩语女声模型
语音文件后处理
图像人脸融合
自动进行人脸区域提取&对齐,并完成面部特征提取,无需额外预处理。
引入3D重建网络对脸型进行拟合迁移,使得融合后的脸型相似度更高。
人脸人体
视觉编辑
DDColor图像上色,相比Deoldify等之前方法在色彩丰富度和语义贴合上大幅提升。
VFI-RAFT视频插帧,和其它SOTA模型相比,在大运动和重复纹理场景下有较好的插帧效果。
DUT-RAFT视频稳像,对多种视频抖动都有稳定的去抖效果,相比原生DUT,能够更好地保持视频清晰度。
底层视觉
Nothing published for this version
Add code generation and code translation from ZHIPU #33
Nothing published for this version
Your coding agent can read these notes before it upgrades. Set up the MCP server →