LLMRepos·Model Development LLM projects

Updated daily

Browse 452 open-source model development LLM projects across Training Toolkits, Fine-Tuning (LoRA / PEFT), Dataset Engineering. Compare stars, growth, languages, licenses, and repository activity.

All categories

Model Development

Training, fine-tuning and research.

452
Repositories
9
Subcategories

Subcategories

Top Model Development repositories

Ranked by current GitHub stars from the latest LLMRepos snapshot.

Showing 40 of 452

Implement a ChatGPT-like LLM in PyTorch from scratch, step by step

Jupyter NotebookOther+840 stars in 7dupdated today
74.6k
stars

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

PythonApache License 2.0+1.4k stars in 7dupdated today

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

PythonApache License 2.0+151 stars in 7dupdated 4d ago

《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程

Jupyter NotebookApache License 2.0+99 stars in 7dupdated 26d ago
21.6k
stars

🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.

PythonApache License 2.0+32 stars in 7dupdated today
19.4k
stars

Toolkit for linearizing PDFs for LLM datasets/training

PythonApache License 2.0+38 stars in 7dupdated 152d ago

中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)

PythonApache License 2.0-1 stars in 7dupdated 128d ago

Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services

Jupyter NotebookMIT License+5 stars in 7dupdated 97d ago
18.3k
stars

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

PythonApache License 2.0+157 stars in 7dupdated today
18.2k
stars

🚀 One-stop solution for creating your AI twin from chat history 💡 Fine-tune LLMs with your chat logs to capture your unique style, then bind to a chatbot to bring your digital self to life.

PythonGNU Affero General Public License v3.0+31 stars in 7dupdated 6d ago
15.3k
stars

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).

PythonApache License 2.0+92 stars in 7dupdated today

Llama中文社区,实时汇总最新Llama学习资料,构建最好的中文Llama大模型开源生态,完全开源可商用

Python-7 stars in 7dupdated 505d ago

Go ahead and axolotl questions

PythonApache License 2.0+29 stars in 7dupdated 3d ago
12.2k
stars

The official GitHub page for the survey paper "A Survey of Large Language Models".

Python+0 stars in 7dupdated 531d ago

Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepower. Maintained by Orchestra Research.

TeXMIT License+220 stars in 7dupdated 70d ago
11k
stars

QLoRA: Efficient Finetuning of Quantized LLMs

Jupyter NotebookMIT License+7 stars in 7dupdated 805d ago
10.7k
stars

Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!

PythonApache License 2.0+63 stars in 7dupdated today
9.9k
stars

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

PythonApache License 2.0+26 stars in 7dupdated 11d ago
9.2k
stars

Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.

PythonOther+148 stars in 7dupdated 8d ago

Lean, fine tuned Opencode multi agent suite · Mix any models · Auto delegate tasks

TypeScriptMIT License+198 stars in 7dupdated today
7.5k
stars

Using Low-rank adaptation to quickly fine-tune diffusion models.

Jupyter NotebookApache License 2.0+3 stars in 7dupdated 885d ago
7.3k
stars

Official release of InternLM series (InternLM, InternLM2, InternLM2.5, InternLM3).

PythonApache License 2.0+4 stars in 7dupdated 299d ago

中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models)

PythonApache License 2.0-4 stars in 7dupdated 128d ago

Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型

Python+2 stars in 7dupdated 670d ago

Efficient Triton Kernels for LLM Training

PythonBSD 2-Clause "Simplified" License+14 stars in 7dupdated today
5.8k
stars

Democratizing Reinforcement Learning for LLMs

PythonApache License 2.0+12 stars in 7dupdated today

A PyTorch Library for Accelerating 3D Deep Learning Research

PythonApache License 2.0-2 stars in 7dupdated today

A curated list of Large Language Model resources, covering model training, serving, fine-tuning, and building LLM applications.

MIT License+6 stars in 7dupdated 371d ago

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

PythonOther+5 stars in 7dupdated today

🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.

PythonOther+103 stars in 7dupdated 1d ago
3.8k
stars

Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs

PythonApache License 2.0+3 stars in 7dupdated 88d ago

DeepResearchAgent is a hierarchical multi-agent system designed not only for deep research tasks but also for general-purpose task solving. The framework leverages a top-level planning agent to coordinate multiple specialized lower-level agents, enabling automated task decomposition and efficient execution across diverse and complex domains.

PythonMIT License+11 stars in 7dupdated 112d ago

🦞 Just talk to your agent — it learns and EVOLVES 🧬.

PythonMIT License+11 stars in 7dupdated 78d ago

总结Prompt&LLM论文,开源数据&模型,AIGC应用

+3 stars in 7dupdated 111d ago

🦖 𝗟𝗲𝗮𝗿𝗻 about 𝗟𝗟𝗠𝘀, 𝗟𝗟𝗠𝗢𝗽𝘀, and 𝘃𝗲𝗰𝘁𝗼𝗿 𝗗𝗕𝘀 for free by designing, training, and deploying a real-time financial advisor LLM system ~ 𝘴𝘰𝘶𝘳𝘤𝘦 𝘤𝘰𝘥𝘦 + 𝘷𝘪𝘥𝘦𝘰 & 𝘳𝘦𝘢𝘥𝘪𝘯𝘨 𝘮𝘢𝘵𝘦𝘳𝘪𝘢𝘭𝘴

Jupyter NotebookMIT License-2 stars in 7dupdated 623d ago
3.4k
stars

A quick guide (especially) for trending instruction finetuning datasets

MIT License-1 stars in 7dupdated 1,000d ago

A framework for prompt tuning using Intent-based Prompt Calibration

PythonApache License 2.0+9 stars in 7dupdated 265d ago
3k
stars

An unofficial PyTorch implementation of the audio LM VALL-E

PythonMIT License-2 stars in 7dupdated 1,202d ago