训练与微调
用自己的数据训练或微调模型的框架与工具
数据截至 8/1 18:48(构建快照,正在获取最新)
按仍在维护排序:先筛掉近 30 天没有提交的项目,再按热度排。 高 star 但早已停更的项目不会出现在这里。
- 1
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
- 2
AirLLM 70B inference with single 4GB GPU
- 3
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
- 4
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
- 5
Go ahead and axolotl questions
- 6
Low-code framework for building custom LLMs, neural networks, and other AI models
- 7
Intelligent Mixture-of-Models Router for Efficient Heterogeneous LLMs Inference
- 8
Build, Evaluate, and Optimize AI Systems. Includes evals, RAG, agents, fine-tuning, synthetic data generation, dataset management, MCP, and more.
- 9
OpenDeepWiki is the open-source version of the DeepWiki project, aiming to provide a powerful knowledge management and collaboration platform. The project is mainly developed using C# and TypeScript, supporting modular design, and is easy to expand and customize.
- 10
OneTrainer is a one-stop solution for all your Diffusion training needs.
- 11
streamline the fine-tuning process for multimodal models: PaliGemma 2, Florence-2, and Qwen2.5-VL
- 12
A simple, performant, and scalable Jax LLM!
- 13
Vendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernetes, and bare metal.
- 14
Distributed AI Model Training and LLM Fine-Tuning on Kubernetes
另有 12 个项目因近期无提交或缺少数据未列入。
训练与微调要落到业务里,还差什么?
开源项目给的是能力,不是方案。数据怎么接、权限怎么管、上线后谁维护,这些才是落地的真正成本。