🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
LLMRepos·Training Toolkits LLM projects
Updated dailyBrowse 275 open-source training toolkits projects in Model Development. Compare GitHub stars, recent growth, languages, licenses, and repository activity.
Training Toolkits
Browse 275 open-source training toolkits projects in Model Development. Compare GitHub stars, recent growth, languages, licenses, and repository activity.
Top Training Toolkits repositories
Ranked by current GitHub stars from the latest LLMRepos snapshot.
Showing 40 of 275
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
End-to-End Speech Processing Toolkit
Efficient Triton Kernels for LLM Training
Democratizing Reinforcement Learning for LLMs
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
Official Repository for QuantHarness
Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
A simple, performant, and scalable Jax LLM!
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
DeepMind's Tacotron-2 Tensorflow implementation
verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"
PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo https://lifeiteng.github.io/valle/index.html
Controllable and fast Text-to-Speech for over 7000 languages!
NX-AI/xlstm
Official repository of the xLSTM.
WaveRNN Vocoder + TTS
Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models
Scalable RL solution for advanced reasoning of language models
Automated Machine Learning on Kubernetes
All-in-one training for vision models (YOLO, ViTs, RT-DETR, DINOv3): pretraining, fine-tuning, distillation.
Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch
Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning
Toolkit for efficient experimentation with Speech Recognition, Text2Speech and NLP
Model compression toolkit engineered for enhanced usability, comprehensiveness, and efficiency.
Recipes to train reward model for RLHF.
Ongoing research training transformer language models at scale, including: BERT & GPT-2
This is now the official location of the Merlin project.
Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.
Tencent Pre-training framework in PyTorch & Pre-trained Model Zoo
Toni-SM/skrl
Modular Reinforcement Learning (RL) library (implemented in PyTorch, JAX, and NVIDIA Warp) with support for Gymnasium/Gym, NVIDIA Isaac Lab, MuJoCo Playground and other environments
[TPAMI 2023] LibFewShot: A Comprehensive Library for Few-shot Learning.
AxisRL is an agentic RL post-training framework built on SGLang rollout, Megatron training, and real-world agent workflows.
DeepLab2 is a TensorFlow library for deep labeling, aiming to provide a unified and state-of-the-art TensorFlow codebase for dense pixel labeling tasks.
Byted PyTorch Distributed for Hyperscale Training of LLMs and RLs
[ACL2023] We introduce LLM-Blender, an innovative ensembling framework to attain consistently superior performance by leveraging the diverse strengths of multiple open-source LLMs. LLM-Blender cut the weaknesses through ranking and integrate the strengths through fusing generation to enhance the capability of LLMs.
🚀 Ultra Recipe for Training Long-Horizon Search Agents - matching frontier AI's search capability with a 20B model + stateful harness