STACKQUADRANT

awesome-ai-apps

rohitg00/awesome-ai-apps
5.0

A curated collection of awesome AI Agents and LLM Apps built with multiple tech stacks, showcasing real-world implementations using OpenAI, Gemini, local models, and various AI frameworks.

LLM Frameworks
804172HTMLApache-2.05mo ago

blades

go-kratos/blades
6.2

Blades is a Go-based multimodal AI Agent framework.

Agent Frameworks
80199GoMIT6d ago

RAG-FiT

IntelLabs/RAG-FiT
5.6

Framework for enhancing LLMs for RAG tasks using fine-tuning.

Fine-tuning Tools
76862PythonApache-2.01mo ago

Trace

microsoft/Trace
6.0

End-to-end Generative Optimization for AI Agents

Prompt Engineering
74959PythonMIT1mo ago

awesome-evals

benchflow-ai/awesome-evals
4.6

A curated, non-BS library of the best resources for building and evaluating AI agents — papers, blogs, talks, tools, benchmarks. Maintained by BenchFlow.

Evaluation & Testing
74365NOASSERTION19d ago

LLM-PowerHouse-A-Curated-Guide-for-Large-Language-Models-with-Custom-Training-and-Inferencing

ghimiresunil/LLM-PowerHouse-A-Curated-Guide-for-Large-Language-Models-with-Custom-Training-and-Inferencing
5.0

LLM-PowerHouse: Unleash LLMs' potential through curated tutorials, best practices, and ready-to-use code for custom training and inferencing.

Fine-tuning Tools
730121Jupyter NotebookMIT4mo ago

voice-ai

rapidaai/voice-ai
6.3

Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channel integration, agent state management, and observability.

Agent Frameworks
698110GoNOASSERTIONtoday

ServerlessLLM

ServerlessLLM/ServerlessLLM
5.7

Serverless LLM Serving for Everyone.

Model Serving
69375PythonApache-2.02mo ago

timber

kossisoroyce/timber
5.3

Ollama for classical ML models. AOT compiler that turns XGBoost, LightGBM, scikit-learn, CatBoost & ONNX models into native C99 inference code. One command to load, one command to serve. 336x faster than Python inference.

Model Serving
68823PythonNOASSERTION3mo ago

long-context-attention

feifeibear/long-context-attention
5.7

USP: Unified (a.k.a. Hybrid, 2D) Sequence Parallel Attention for Long Context Transformers Model Training and Inference

Fine-tuning Tools
68181PythonApache-2.02mo ago

gollm

teilomillet/gollm
5.8

Unified Go interface for Language Model (LLM) providers. Simplifies LLM integration with flexible prompt management and common task functions.

Prompt Engineering
67166GoApache-2.04mo ago

Awesome-LLM-Eval

onejune2018/Awesome-LLM-Eval
4.6

Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界.

Evaluation & Testing
65281MIT7mo ago

aimock

CopilotKit/aimock
6.4

Mock everything your AI app talks to — LLM APIs, MCP, A2A, vector DBs, search. One package, one port, zero dependencies.

Evaluation & Testing
65146TypeScriptMITtoday

Awesome-LLM-in-Social-Science

ValueByte-AI/Awesome-LLM-in-Social-Science
5.0

Awesome papers involving LLMs in Social Science.

Evaluation & Testing
63649MIT1mo ago

LLMTornado

lofcz/LLMTornado
6.6

The .NET library to build AI agents with 30+ built-in connectors.

Agent Frameworks
627108C#MITtoday

agent-skills-eval

darkrishabh/agent-skills-eval
5.2

A test runner for agentskills.io-style AI agent skills

Evaluation & Testing
62534TypeScriptMIT5d ago

gitagent

open-gitagent/gitagent
5.8

A framework-agnostic, git-native standard for defining AI agents

Agent Frameworks
615117TypeScriptMIT8d ago

daydreams

daydreamsai/daydreams
5.8

Daydreams is a set of tools for building agents for commerce

Agent Frameworks
611135TypeScriptMIT4mo ago

fastapi-ml-skeleton

eightBEC/fastapi-ml-skeleton
4.5

FastAPI Skeleton App to serve machine learning models production-ready.

Model Serving
60490PythonApache-2.06mo ago

yalm

andrewkchan/yalm
3.7

Yet Another Language Model: LLM inference in C++/CUDA, no libraries except for I/O

Inference Engines
59164C++10mo ago

ICLR2025-Papers-with-Code

yinizhilian/ICLR2025-Papers-with-Code
3.3

历年ICLR论文和开源项目合集,包含ICLR2021、ICLR2022、ICLR2023、ICLR2024、ICLR2025.

Fine-tuning Tools
589321y ago

tessera

zengxiao-he/tessera
4.5

From teacher to tiles — a from-scratch LLM distillation & serving engine: custom Triton/CUDA kernels, FSDP distillation, paged-KV continuous batching, speculative decoding, a Rust gateway, a JAX oracle, and interpretability tooling.

Inference Engines
5889PythonNOASSERTION1mo ago

LLM-FineTuning-Large-Language-Models

rohan-paul/LLM-FineTuning-Large-Language-Models
3.6

LLM (Large Language Model) FineTuning

Fine-tuning Tools
576139Jupyter Notebook1y ago

langtest

Pacific-AI-Corp/langtest
5.9

Deliver safe & effective language models

Evaluation & Testing
56351PythonApache-2.0today

langtest

PacificAI/langtest
5.9

Deliver safe & effective language models

Evaluation & Testing
56351PythonApache-2.0today

openinfer

openinfer-project/openinfer
6.1

Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2

Inference Engines
55885RustApache-2.0today

KuiperLLama

zjhellofss/KuiperLLama
4.0

校招、秋招、春招、实习好项目,带你从零动手实现支持LLama2/3和Qwen2.5的大模型推理框架。

Inference Engines
552143C++8mo ago

awesome-on-policy-distillation

chrisliu298/awesome-on-policy-distillation
4.4

A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models

Fine-tuning Tools
54420CC0-1.0today

pinferencia

underneathall/pinferencia
4.7

Python + Inference - Model Deployment library in Python. Simplest model inference server ever.

Model Serving
54383PythonApache-2.03y ago

Athena-Public

winstonkoh87/Athena-Public
6.0

The Linux OS for AI Agents — Persistent memory, autonomy, and time-awareness for any LLM. Own the state. Rent the intelligence.

LLM Frameworks
54172PythonMITtoday

TurboOCR

aiptimizer/TurboOCR
5.6

Fast GPU OCR server. 270 img/s on FUNSD. TensorRT FP16, PP-OCRv5, HTTP + gRPC.

Model Serving
52856C++MIT1d ago

continuous-eval

relari-ai/continuous-eval
4.7

Data-Driven Evaluation for LLM-Powered Applications

Evaluation & Testing
51638PythonApache-2.01y ago

LLM-VM

anarchy-ai/LLM-VM
4.8

irresponsible innovation. Try now at https://chat.dev/

Fine-tuning Tools
491137PythonMIT2y ago

agency

operand/agency
5.0

A fast and minimal framework for building agentic systems

Agent Frameworks
48728PythonMIT1mo ago

fakecloud

faiscadev/fakecloud
5.8

Free, open-source AWS emulator. LocalStack alternative: 26 services, 1,924 operations, 100% conformance. No account, no auth token, no paid tier.

Evaluation & Testing
48733RustAGPL-3.0today

ome

ome-projects/ome
6.1

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

Model Serving
48186GoApache-2.0today

Finetune_LLMs

mallorbc/Finetune_LLMs
3.8

Repo for fine-tuning Casual LLMs

Fine-tuning Tools
46786PythonAGPL-3.02y ago

awsome-distributed-training

awslabs/awsome-distributed-training
5.6

Collection of best practices, reference architectures, model training examples and utilities to train large models on AWS.

Fine-tuning Tools
465199ShellMIT-0today

agentsilex

howl-anderson/agentsilex
4.7

A transparent, minimal, and hackable agent framework. ~300 lines of readable code. Full control, no magic.

Agent Frameworks
45245PythonMIT6mo ago

JetStream

AI-Hypercomputer/JetStream
4.9

JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs welcome).

Model Serving
45167PythonApache-2.06mo ago

Aquila2

FlagAI-Open/Aquila2
3.6

The official repo of Aquila2 series proposed by BAAI, including pretrained & chat large language models.

Fine-tuning Tools
44632Python1y ago

xFasterTransformer

intel/xFasterTransformer
4.3

xFasterTransformer — open-source AI/LLM project.

Model Serving
43576C++Apache-2.010mo ago

gpu-rest-engine

NVIDIA/gpu-rest-engine
3.7

A REST API for Caffe using Docker and Go

Model Serving
42293C++BSD-3-Clause8y ago

InternEvo

InternLM/InternEvo
4.8

InternEvo is an open-sourced lightweight training framework aims to support model pre-training without the need for extensive dependencies.

Fine-tuning Tools
42167PythonApache-2.011mo ago

Awesome-LLM-Prompt-Optimization

jxzhangjhu/Awesome-LLM-Prompt-Optimization
4.2

Awesome-LLM-Prompt-Optimization: a curated list of advanced prompt optimization and tuning methods in Large Language Models

Prompt Engineering
4122325d ago

awesome-azure-openai-llm

kimtth/awesome-azure-openai-llm
4.6

A curated collection of resources for 🌌 Azure OpenAI, 🦙 LLMs (RAG, Agents).

Agent Frameworks
40458Python4d ago

tiger

tigerlab-ai/tiger
4.3

Open Source LLM toolkit to build trustworthy LLM applications. TigerArmor (AI safety), TigerRAG (embedding, RAG), TigerTune (fine-tuning)

Fine-tuning Tools
40327Jupyter NotebookApache-2.02y ago

LightRFT

opendilab/LightRFT
5.0

LightRFT: Light, Efficient, Omni-modal & Reward-model Driven Reinforcement Fine-Tuning Framework

Fine-tuning Tools
40211PythonApache-2.02mo ago