LLMRepos·Training Toolkits LLM projects

Updated daily

Browse 275 open-source training toolkits projects in Model Development. Compare GitHub stars, recent growth, languages, licenses, and repository activity.

Training Toolkits

Browse 275 open-source training toolkits projects in Model Development. Compare GitHub stars, recent growth, languages, licenses, and repository activity.

275
Repositories
452
Model Development

Top Training Toolkits repositories

Ranked by current GitHub stars from the latest LLMRepos snapshot.

Showing 40 of 275

45.9k
stars

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

PythonMozilla Public License 2.0+39 stars in 7dupdated 738d ago

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

PythonApache License 2.0+97 stars in 7dupdated 91d ago
18.3k
stars

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

PythonApache License 2.0+157 stars in 7dupdated today
10.3k
stars

Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.

PythonMIT License+21 stars in 7dupdated 152d ago
10.2k
stars

:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)

Jupyter NotebookMozilla Public License 2.0+3 stars in 7dupdated 1,019d ago

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

PythonApache License 2.0+26 stars in 7dupdated 11d ago
9.9k
stars

End-to-End Speech Processing Toolkit

PythonApache License 2.0+10 stars in 7dupdated today

Efficient Triton Kernels for LLM Training

PythonBSD 2-Clause "Simplified" License+14 stars in 7dupdated today
5.8k
stars

Democratizing Reinforcement Learning for LLMs

PythonApache License 2.0+12 stars in 7dupdated today

Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.

PythonApache License 2.0+796 stars in 7dupdated today
2.8k
stars

Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics

PythonMIT Licenseupdated 1d ago

A simple, performant, and scalable Jax LLM!

PythonApache License 2.0+14 stars in 7dupdated today
2.4k
stars

HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

PythonMIT License+1 stars in 7dupdated 758d ago

DeepMind's Tacotron-2 Tensorflow implementation

PythonMIT License-1 stars in 7dupdated 1,145d ago

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

PythonApache License 2.0+24 stars in 7dupdated 76d ago
2.2k
stars

PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo https://lifeiteng.github.io/valle/index.html

PythonApache License 2.0+0 stars in 7dupdated 348d ago

Controllable and fast Text-to-Speech for over 7000 languages!

PythonApache License 2.0-1 stars in 7dupdated 211d ago
2.2k
stars

Official repository of the xLSTM.

PythonApache License 2.0-1 stars in 7dupdated 88d ago
2.2k
stars

WaveRNN Vocoder + TTS

PythonMIT License-1 stars in 7dupdated 1,514d ago

Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models

Jupyter NotebookApache License 2.0+44 stars in 7dupdated 3d ago
1.9k
stars

Scalable RL solution for advanced reasoning of language models

PythonApache License 2.0+2 stars in 7dupdated 524d ago
1.7k
stars

Automated Machine Learning on Kubernetes

PythonApache License 2.0+2 stars in 7dupdated 18d ago

All-in-one training for vision models (YOLO, ViTs, RT-DETR, DINOv3): pretraining, fine-tuning, distillation.

PythonGNU Affero General Public License v3.0+5 stars in 7dupdated today

Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch

Jupyter NotebookMIT License+0 stars in 7dupdated 855d ago
1.6k
stars

Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning

PythonMIT License+15 stars in 7dupdated today
1.6k
stars

Toolkit for efficient experimentation with Speech Recognition, Text2Speech and NLP

PythonApache License 2.0+0 stars in 7dupdated 1,931d ago
1.5k
stars

Model compression toolkit engineered for enhanced usability, comprehensiveness, and efficiency.

PythonOther+11 stars in 7dupdated 17d ago

This is now the official location of the Merlin project.

PythonApache License 2.0+0 stars in 7dupdated 2,365d ago
1.2k
stars

Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.

Python+95 stars in 7dupdated 95d ago

Tencent Pre-training framework in PyTorch & Pre-trained Model Zoo

PythonOther+1 stars in 7dupdated 750d ago
1.1k
stars

Modular Reinforcement Learning (RL) library (implemented in PyTorch, JAX, and NVIDIA Warp) with support for Gymnasium/Gym, NVIDIA Isaac Lab, MuJoCo Playground and other environments

PythonMIT License+0 stars in 7dupdated 105d ago
1.1k
stars

[TPAMI 2023] LibFewShot: A Comprehensive Library for Few-shot Learning.

PythonMIT License-1 stars in 7dupdated 301d ago
1.1k
stars

AxisRL is an agentic RL post-training framework built on SGLang rollout, Megatron training, and real-world agent workflows.

PythonApache License 2.0+4 stars in 7dupdated 21d ago

DeepLab2 is a TensorFlow library for deep labeling, aiming to provide a unified and state-of-the-art TensorFlow codebase for dense pixel labeling tasks.

PythonApache License 2.0+0 stars in 7dupdated 1,225d ago

Byted PyTorch Distributed for Hyperscale Training of LLMs and RLs

PythonApache License 2.0+1 stars in 7dupdated 174d ago

[ACL2023] We introduce LLM-Blender, an innovative ensembling framework to attain consistently superior performance by leveraging the diverse strengths of multiple open-source LLMs. LLM-Blender cut the weaknesses through ranking and integrate the strengths through fusing generation to enhance the capability of LLMs.

PythonApache License 2.0+0 stars in 7dupdated 671d ago
985
stars

🚀 Ultra Recipe for Training Long-Horizon Search Agents - matching frontier AI's search capability with a 20B model + stateful harness

PythonApache License 2.0+2 stars in 7dupdated 70d ago