STACKQUADRANT

Tracely-ai

Jwuthri/Tracely-ai
5.8

Trace-native CI/CD for AI agents — production failures become regression tests that block the PR. Auto-detect, cluster, freeze into hermetic cases, replay in CI for $0.

Evaluation & Testing
1.2k104PythonMITtoday

cashclaw

moltlaunch/cashclaw
5.2

An autonomous agent that takes work, does work, gets paid, and gets better at it.

Agent Frameworks
1.2k222TypeScriptMIT5mo ago

kvcached

ovg-project/kvcached
5.8

Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond

Inference Engines
1.1k133PythonApache-2.05d ago

DataDreamer

datadreamer-dev/DataDreamer
4.9

DataDreamer: Prompt. Generate Synthetic Data. Train & Align Models.   🤖💤

Fine-tuning Tools
1.1k58PythonMIT1y ago

nobodywho

nobodywho-ooo/nobodywho
6.6

NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.

Inference Engines
1.1k77RustEUPL-1.2today

judgeval

JudgmentLabs/judgeval
6.6

The open source post-building layer for agents. Our environment data and evals power agent post-training (RL, SFT) and monitoring.

Evaluation & Testing
1.1k96PythonApache-2.04d ago

SLAM-LLM

X-LANCE/SLAM-LLM
5.4

A Framework for Speech, Language, Audio, Music Processing with Large Language Model

LLM Frameworks
1.1k117PythonMIT7mo ago

pydantic-deepagents

vstorm-co/pydantic-deepagents
6.6

Python Deep Agent framework built on top of Pydantic-AI, designed to help you quickly build production-grade autonomous AI agents with planning, filesystem operations, subagent delegation, skills, and structured outputs—in just 10 lines of code.

Agent Frameworks
1.0k126PythonMIT5d ago

veScale

volcengine/veScale
5.1

Byted PyTorch Distributed for Hyperscale Training of LLMs and RLs

Fine-tuning Tools
1.0k65PythonApache-2.05mo ago

TurboOCR

aiptimizer/TurboOCR
5.9

Fast GPU OCR server. 270 img/s on FUNSD. TensorRT FP16, PP-OCRv5, HTTP + gRPC.

Model Serving
1.0k100C++MIT8d ago

FinSight-AI

juanjuandog/FinSight-AI
5.3

AI equity research agent with resilient workflows, Redis Lua single-flight, pgvector RAG, versioned reports, evidence tracing, and RAG evaluation.

Evaluation & Testing
1.0k61JavaMIT4d ago

WHartTest

MGdaasLab/WHartTest
6.4

WHartTest 是一款AI驱动的测试自动化平台,实现从需求到可执行测试用例的自动化生成与管理,帮助测试团队提升效率与覆盖率。 (WHartTest is an AI-driven test automation platform that automates the generation and management of executable test cases from requirements, helping testing teams improve efficiency and coverage.)

Evaluation & Testing
1.0k159PythonMIT1d ago

awesome-dsh-plugin

Anil-matcha/awesome-dsh-plugin
5.5

A curated list of plugins for DeepSeek Harness (dsh) - DeepSeek Harness plugin ecosystem

Agent Frameworks
9882563d ago

start-llms

louisfb01/start-llms
5.2

A complete guide to start and improve your LLM skills in 2026 with little background in the field and stay up-to-date with the latest news and state-of-the-art techniques!

Fine-tuning Tools
981127MIT7mo ago

Nanoflow

efeslab/Nanoflow
4.6

A throughput-oriented high-performance serving framework for LLMs

Model Serving
97452Jupyter Notebook5mo ago

sglang-omni

sgl-project/sglang-omni
6.9

SGLang-Omni empowers high-performance serving for TTS, ASR, speech and omni models.

Model Serving
968395PythonApache-2.0today

mlxstudio

jjang-ai/mlxstudio
5.4

MLX Studio - Home of JANG_Q - Image Gen/Edit + Chat/Code All in one - + OpenClaw (Anthropic API)

Inference Engines
95965today

scenario

langwatch/scenario
6.0

Agentic testing for agentic codebases

Evaluation & Testing
95578PythonMITtoday

mcp-framework

QuantGeekDev/mcp-framework
6.0

The Typescript MCP Framework

LLM Frameworks
928113TypeScriptMIT4mo ago

AI-Compass

tingaicompass/AI-Compass
5.6

“AI-Compass”将为社区指引在 AI 技术海洋中航行的方向,无论你是初学者还是进阶开发者,都能在这里找到通往 AI 各大方向的路径。旨在帮助开发者系统性地了解 AI 的核心概念、主流技术、前沿趋势,并通过实践掌握从理论到落地的全过程。

Fine-tuning Tools
923123Pythontoday

model_server

openvinotoolkit/model_server
6.5

A scalable inference server for models optimized with OpenVINO™

Model Serving
920271C++Apache-2.0today

ZhiLight

zhihu/ZhiLight
5.1

A highly optimized LLM inference acceleration engine for Llama and its variants.

Inference Engines
908104C++Apache-2.05mo ago

mosec

mosecorg/mosec
6.4

A high-performance ML model serving framework, offers dynamic batching and CPU/GPU pipelines to fully exploit your compute machine

Model Serving
90273PythonApache-2.018d ago

aimock

CopilotKit/aimock
6.6

Mock everything your AI app talks to — LLM APIs, MCP, A2A, vector DBs, search. One package, one port, zero dependencies.

Evaluation & Testing
90264TypeScriptMITtoday

agent-qa

vostride/agent-qa
5.6

The self-improving Agentic QA harness with Memory. Write tests in natural language.
 Catch regressions before releases ship.

Evaluation & Testing
87915TypeScriptNOASSERTION24d ago

Foundry

promptise-com/Foundry
6.0

The foundation layer for agentic intelligence.

Agent Frameworks
869129PythonApache-2.08d ago

DeepMCPAgent

cryxnet/DeepMCPAgent
6.1

Model-agnostic plug-n-play LangChain/LangGraph agents powered entirely by MCP tools over HTTP/SSE.

Agent Frameworks
869129PythonApache-2.08d ago

cerebellum

theredsix/cerebellum
4.9

Browser automation system that uses AI-driven planning to navigate web pages and perform goals.

Fine-tuning Tools
86457PythonMIT2mo ago

pipeless

pipeless-ai/pipeless
5.0

An open-source computer vision framework to build and deploy apps in minutes

Model Serving
85252RustApache-2.02y ago

awesome-evals

benchflow-ai/awesome-evals
5.1

A curated, non-BS library of the best resources for building and evaluating AI agents — papers, blogs, talks, tools, benchmarks. Maintained by BenchFlow.

Evaluation & Testing
84789NOASSERTION7d ago

Yatai

bentoml/Yatai
5.9

Model Deployment at Scale on Kubernetes 🦄️

Model Serving
84076TypeScriptNOASSERTION3mo ago

awesome-ai-apps

rohitg00/awesome-ai-apps
4.9

A curated collection of awesome AI Agents and LLM Apps built with multiple tech stacks, showcasing real-world implementations using OpenAI, Gemini, local models, and various AI frameworks.

LLM Frameworks
823174HTMLApache-2.06mo ago

Adan

sail-sg/Adan
4.4

Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models

Fine-tuning Tools
82270PythonApache-2.01y ago

blades

go-kratos/blades
6.2

Blades is a Go-based multimodal AI Agent framework.

Agent Frameworks
810102GoMIT14d ago

RAG-FiT

IntelLabs/RAG-FiT
5.5

Framework for enhancing LLMs for RAG tasks using fine-tuning.

Fine-tuning Tools
76961PythonApache-2.02mo ago

awesome-on-policy-distillation

chrisliu298/awesome-on-policy-distillation
4.5

A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models

Fine-tuning Tools
75926CC0-1.013d ago

Trace

microsoft/Trace
5.9

End-to-end Generative Optimization for AI Agents

Prompt Engineering
75461PythonMIT2mo ago

LLM-PowerHouse-A-Curated-Guide-for-Large-Language-Models-with-Custom-Training-and-Inferencing

ghimiresunil/LLM-PowerHouse-A-Curated-Guide-for-Large-Language-Models-with-Custom-Training-and-Inferencing
4.9

LLM-PowerHouse: Unleash LLMs' potential through curated tutorials, best practices, and ready-to-use code for custom training and inferencing.

Fine-tuning Tools
731121Jupyter NotebookMIT5mo ago

voice-ai

rapidaai/voice-ai
6.3

Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channel integration, agent state management, and observability.

Agent Frameworks
720121GoNOASSERTIONtoday

ServerlessLLM

ServerlessLLM/ServerlessLLM
5.6

Serverless LLM Serving for Everyone.

Model Serving
71176PythonApache-2.03mo ago

agent-skills-eval

darkrishabh/agent-skills-eval
5.2

A test runner for agentskills.io-style AI agent skills

Evaluation & Testing
70635TypeScriptMIT22d ago

long-context-attention

feifeibear/long-context-attention
5.5

USP: Unified (a.k.a. Hybrid, 2D) Sequence Parallel Attention for Long Context Transformers Model Training and Inference

Fine-tuning Tools
68784PythonApache-2.03mo ago

timber

kossisoroyce/timber
5.2

Ollama for classical ML models. AOT compiler that turns XGBoost, LightGBM, scikit-learn, CatBoost & ONNX models into native C99 inference code. One command to load, one command to serve. 336x faster than Python inference.

Model Serving
68723PythonNOASSERTION4mo ago

gollm

teilomillet/gollm
5.7

Unified Go interface for Language Model (LLM) providers. Simplifies LLM integration with flexible prompt management and common task functions.

Prompt Engineering
67467GoApache-2.05mo ago

gitagent

open-gitagent/gitagent
5.8

A framework-agnostic, git-native standard for defining AI agents

Agent Frameworks
659123RustMIT8d ago

Awesome-LLM-Eval

onejune2018/Awesome-LLM-Eval
4.5

Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界.

Evaluation & Testing
65884MIT9mo ago

openinfer

openinfer-project/openinfer
6.2

Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2

Inference Engines
657103RustApache-2.0today

pegainfer

pegainfer-project/pegainfer
6.5

Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2

Inference Engines
657103RustApache-2.0today