Skip to content
aicoolies logo
ModelScope logo

ms-swift

ModelScope's fine-tuning framework supporting 600+ models

ms-swift is ModelScope's open-source framework for fine-tuning over 600 large language and multimodal models. It supports SFT, DPO, RLHF, LoRA, QLoRA, and full fine-tuning with a web UI and CLI interface. Optimized for the Chinese AI ecosystem with native ModelScope Hub integration alongside Hugging Face support. Over 13,500 GitHub stars.

About ms-swift

MS-SWIFT (ModelScope SWIFT) is a comprehensive fine-tuning and deployment framework supporting 600+ LLMs and 400+ multimodal models from the open-source ecosystem. Built by Alibaba ModelScope, it abstracts away training complexity: whether pre-training from scratch, instruction-tuning on custom data, or applying preference learning (GRPO, DPO), MS-SWIFT provides unified APIs and boilerplate-free code. Supported model families span Qwen, DeepSeek-R1, InternLM, GLM, Mistral, and Llama for text, plus vision models like Qwen-VL, InternVL, and LLaVA, ensuring coverage across the latest open-source innovations.

The framework integrates cutting-edge training techniques: megatron-style parallelism (TP, PP, CP, EP) accelerates training on multi-GPU and multi-node clusters, while GRPO and variants (DAPO, GSPO, SAPO) implement modern reinforcement learning alignment without boilerplate. MS-SWIFT handles data loading, tokenization, optimizer scheduling, checkpoint management, and eval harnesses. For inference, it supports accelerators (vLLM, SGLang, LmDeploy) and exports quantized models (AWQ, GPTQ, FP8). The OpenAI-compatible serving API lets fine-tuned models drop in as LLM backends.

ML teams adopting Qwen or other ModelScope models benefit from batteries-included tooling rather than stitching together Transformers, Accelerate, vLLM, and custom utilities. The focus on GRPO and modern alignment methods reflects enterprise demand for fine-tuned models that stay aligned. Active development tracks emerging model releases, so support for new architectures appears weeks after release. For production deployments using open-source models, MS-SWIFT reduces engineering overhead significantly.

Pricing & Platform Specs

Pricing Summary

Free and 100% open source under the Apache-2.0 license with $0 software licensing fees. ms-swift provides fine-tuning, RLHF alignment, and inference pipelines for 600+ LLMs and 400+ multimodal models; users pay only for their own GPU compute infrastructure.

full pricing breakdown →

Supported Platforms

Python, CUDA GPUs, ModelScope/Hugging Face

Explore categories, tags & use cases

Categories

Unified framework for fine-tuning 100+ large language models

LLaMA-Factory is an open-source toolkit providing a unified interface for fine-tuning over 100 LLMs and vision-language models. It supports SFT, RLHF with PPO and DPO, LoRA and QLoRA for memory-efficient training, and continuous pre-training. The LLaMA Board web UI enables no-code configuration, while CLI and YAML workflows serve advanced users. Integrates with Hugging Face, ModelScope, vLLM, and SGLang for model deployment.

Open Source

Meta's official PyTorch library for LLM fine-tuning

torchtune is Meta's official PyTorch-native library for fine-tuning large language models. It provides composable building blocks for training recipes covering LoRA, QLoRA, full fine-tuning, DPO, and knowledge distillation. Supports Llama, Mistral, Gemma, Qwen, and Phi model families with distributed training across multiple GPUs. Designed as a hackable, dependency-minimal alternative to higher-level frameworks.

freeOpen Source

Side-by-Side Comparisons

ModelScope logo
ms-swift
vs
LLaMA Factory project logo
LLaMA-Factory

ms-swift vs LLaMA-Factory — ModelScope Fine-Tuning Hub vs Universal Training Orchestrator

ms-swift and LLaMA-Factory both simplify LLM fine-tuning with web UIs and CLI interfaces but serve different primary ecosystems. ms-swift by ModelScope supports over 600 models with native integration into China's ModelScope Hub alongside Hugging Face. LLaMA-Factory provides the most popular fine-tuning framework globally with 69,000+ stars, comprehensive training method coverage, and deep Hugging Face ecosystem integration.

Community experience

Sources & verification

Sources checked
Content verified

Verification dates are editorial checks. Routine CMS saves and automatic updatedAt timestamps do not advance them.

FAQ

What is ms-swift?

ms-swift is ModelScope's open-source framework for fine-tuning over 600 large language and multimodal models. It supports SFT, DPO, RLHF, LoRA, QLoRA, and full fine-tuning with a web UI and CLI interface. Optimized for the Chinese AI ecosystem with native ModelScope Hub integration alongside Hugging Face support. Over 13,500 GitHub stars.

Is ms-swift free?

Yes — ms-swift is open source and free to use. Free and 100% open source under the Apache-2.0 license with $0 software licensing fees. ms-swift provides fine-tuning, RLHF alignment, and inference pipelines for 600+ LLMs and 400+ multimodal models; users pay only for their own GPU compute infrastructure.

Is ms-swift open source?

Yes — ms-swift is open source.

Is ms-swift still maintained?

Yes — ms-swift is active. Its listing was last verified on September 6, 2026.

What are the best ms-swift alternatives?

The first editor-selected ms-swift alternatives are LLaMA-Factory, torchtune.