Skip to main content

Category · 26 repos

Fine-tuning & Training

Tools for training, fine-tuning and aligning models, from LoRA to reinforcement learning. Ranked by star velocity over the last 24 hours.

1
overmind-core/overmind

The platform for continuously improving AI agents.

PythonFine-tuning & TrainingAI Agents
2
bandr-ai/bandits

Trace mining for post-training. Filter trajectories, draft verifiers, export SFT and RL data.

PythonFine-tuning & Training
3
TokenRhythm/NeoHorse

NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness.

PythonFine-tuning & TrainingAI Agents
4
calmrocks/ai-engineer-notebooks

Hands-on, framework-free Colab notebooks for the AI Engineer / Forward Deployed Engineer (FDE) skill set — model APIs, structured output, tool calling, RAG, evals-as-the-spine, agents (loop from scratch, tool design, guardrails, MCP, Skills), fine-tuning vs LoRA, prompt-injection/security, LLMOps, a

Jupyter NotebookFine-tuning & TrainingLLMOps & Gateways
5
OpenDCAI/DataFlex

[NeurIPS 2026] DataFlex: A Unified Benchmark and Evaluation Platform for Data-Centric Training of Large Language Models

PythonFine-tuning & TrainingMachine Learning & Data Science
6
firelex/jeff

Millisecond decisions, any domain: a 0.8B open "System 1" model that picks between your options with calibrated probabilities. One base, swappable LoRA adapters, on your own hardware.

PythonFine-tuning & Training
7
Blaizzy/mlx-vlm

MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.

PythonFine-tuning & TrainingLocal LLMs
8
aisa-group/PostTrainBench

Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours

PythonFine-tuning & TrainingAI Coding Assistants
9
iamjrmh/usersync

Word-accurate .lrc lyric sync. Paste your lyrics, point it at any audio file, get back an enhanced LRC with per-word timestamps. Powered by whisper.cpp + WhisperX (wav2vec2 forced alignment) + Demucs vocal isolation. Prebuilt for NVIDIA CUDA, Vulkan, and CPU. Manual editor for fine-tuning, live prev

C++Fine-tuning & TrainingAI Safety & Red Teaming
10
whitecircle/halo

Halo is an open-source framework built by White Circle for training large language and multimodal models

PythonFine-tuning & TrainingComputer Vision
11
RegiaYoung/SlideDP

Scaling Host-Resident LLM Fine-Tuning Across Multiple GPUs

PythonFine-tuning & Training
12
alexiglad/XM

PyTorch Code for Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation

PythonFine-tuning & TrainingTesting & QA
13
medhiclb/HelixAI

Your own AI workspace on your own machines: chat, agents, coding, knowledge bases (RAG) and fine-tuning with local models. Open source (AGPL-3.0).

TypeScriptFine-tuning & TrainingRAG & Knowledge Bases
14
shootthesound/Fizgig

Fizgig — LoRA & Fine-tune Studio for Klein 9B, Krea 2, MiniMax H3 & Qwen Image 2.1: train, fine-tune, profile, repair and extract LoRAs & LoKRs

PythonFine-tuning & TrainingImage Generation & Editing
15
netease-youdao/Confucius4-TTS

Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine

PythonFine-tuning & TrainingText to Speech & Voice
16
MakazhanAlpamys/Soup

Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.

PythonFine-tuning & TrainingLocal LLMs
17
hiyouga/LlamaFactory

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

PythonFine-tuning & TrainingLLM Inference & Serving
18
LMIS-ORG/slime-agentic

A project implementing various agentic RL based on the Slime post-training framework

PythonFine-tuning & TrainingAI Agents
19
bubbliiiing/segformer-pytorch

Source code for segformer-pytorch, which can be used to train your own models.

PythonFine-tuning & Training
20
Human-Agent-Society/reef

Infrastructure for continually self‑improving agents

PythonFine-tuning & TrainingAI Agents
21
linkedin/Liger-Kernel

Efficient Triton Kernels for LLM Training

PythonFine-tuning & TrainingLLM Inference & Serving
22
sapientinc/HRM-Text

HRM-Text is a 1B text generation model based on the HRM architecture, strengthened by task completion and latent space reasoning.

PythonFine-tuning & Training
23
whoashish115/kitsune-tales-qwen

Kitsune-Tales-E4B-JP and -EN: LoRA fine-tunes of Gemma 4 E4B that write original fantasy light-novel fiction in Japanese and English. Synthetic data, SFT + DPO, bootstrap-CI evaluation, a validated LLM judge, safety audit, GGUF builds.

PythonFine-tuning & TrainingLocal LLMs
24
POInfra-AI/POInfra

Policy Optimization for Generative Models: diffusion and flow policies, online fine-tuning, and inference-time guidance and planning.

Fine-tuning & Training
25
flyteorg/flyte

Type-safe, distributed orchestration of agents, ML pipelines, and real-time inference on your k8s — in pure Python with async/await, also other languages (rust, go and ts)

GoFine-tuning & TrainingMachine Learning & Data Science
26
axolotl-ai-cloud/axolotl

Go ahead and axolotl questions

PythonFine-tuning & Training