STACKQUADRANT

Fine-tuning Tools

Tools for fine-tuning, training, and adapting foundation models

58 repos

beam-cloud/beta9

6.9

Ultrafast serverless GPU inference, sandboxes, and background jobs

1.8k162Go

NVIDIA-NeMo/Curator

6.4

Scalable data pre processing and curation toolkit for LLMs

1.7k321Python

bespokelabsai/curator

6.4

Synthetic data curation for post-training and structured data extraction

1.7k146Python

intelligent-machine-learning/dlrover

6.5

DLRover: An Automatic Distributed Deep Learning System

1.7k218Python

R6410418/Jackrong-llm-finetuning-guide

5.7

Jackrong-llm-finetuning-guide — open-source AI/LLM project.

1.7k268Jupyter Notebook

utkuozdemir/nvidia_gpu_exporter

6.6

Nvidia GPU exporter for prometheus using nvidia-smi binary

1.5k154Go

ARahim3/mlx-tune

5.4

Bringing the Unsloth experience to Mac users via Apple's MLX framework

1.4k91Python

AutoArk/TinyEngram

4.7

Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.

1.3k90Python

SakanaAI/text-to-lora

5.0

Hypernetworks that adapt LLMs for specific benchmark tasks using only textual task description as the input

1.3k90Python

InternScience/GraphGen

6.3

GraphGen: Enhancing Supervised Fine-Tuning for LLMs with Knowledge-Driven Synthetic Data Generation

1.2k97Python

datadreamer-dev/DataDreamer

4.9

DataDreamer: Prompt. Generate Synthetic Data. Train & Align Models.   🤖💤

1.1k58Python

volcengine/veScale

5.1

Byted PyTorch Distributed for Hyperscale Training of LLMs and RLs

1.0k65Python

louisfb01/start-llms

5.2

A complete guide to start and improve your LLM skills in 2026 with little background in the field and stay up-to-date with the latest news and state-of-the-art techniques!

981127

tingaicompass/AI-Compass

5.6

“AI-Compass”将为社区指引在 AI 技术海洋中航行的方向,无论你是初学者还是进阶开发者,都能在这里找到通往 AI 各大方向的路径。旨在帮助开发者系统性地了解 AI 的核心概念、主流技术、前沿趋势,并通过实践掌握从理论到落地的全过程。

930124Python

theredsix/cerebellum

4.9

Browser automation system that uses AI-driven planning to navigate web pages and perform goals.

86457Python

sail-sg/Adan

4.4

Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models

82270Python

IntelLabs/RAG-FiT

5.5

Framework for enhancing LLMs for RAG tasks using fine-tuning.

76861Python

chrisliu298/awesome-on-policy-distillation

4.6

A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models

76827

ghimiresunil/LLM-PowerHouse-A-Curated-Guide-for-Large-Language-Models-with-Custom-Training-and-Inferencing

4.9

LLM-PowerHouse: Unleash LLMs' potential through curated tutorials, best practices, and ready-to-use code for custom training and inferencing.

731121Jupyter Notebook

feifeibear/long-context-attention

5.5

USP: Unified (a.k.a. Hybrid, 2D) Sequence Parallel Attention for Long Context Transformers Model Training and Inference

69084Python

yinizhilian/ICLR2025-Papers-with-Code

3.3

历年ICLR论文和开源项目合集,包含ICLR2021、ICLR2022、ICLR2023、ICLR2024、ICLR2025.

59134

rohan-paul/LLM-FineTuning-Large-Language-Models

3.6

LLM (Large Language Model) FineTuning

578136Jupyter Notebook

anarchy-ai/LLM-VM

4.7

irresponsible innovation. Try now at https://chat.dev/

490139Python

awslabs/awsome-distributed-training

5.9

Collection of best practices, reference architectures, model training examples and utilities to train large models on AWS.

473206Shell