STACKQUADRANT

Awesome-LLM-in-Social-Science

ValueByte-AI/Awesome-LLM-in-Social-Science
5.3

Awesome papers involving LLMs in Social Science.

Evaluation & Testing
64652MIT9d ago

LLMTornado

lofcz/LLMTornado
6.6

The .NET library to build AI agents with 30+ built-in connectors.

Agent Frameworks
638112C#MIT10d ago

daydreams

daydreamsai/daydreams
5.6

Daydreams is a set of tools for building agents for commerce

Agent Frameworks
616133TypeScriptMIT5mo ago

fastapi-ml-skeleton

eightBEC/fastapi-ml-skeleton
4.4

FastAPI Skeleton App to serve machine learning models production-ready.

Model Serving
60493PythonApache-2.07mo ago

ponytail-improved

0xwilliamortiz/ponytail-improved
5.6

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

Prompt Engineering
601129JavaScriptMIT23d ago

yalm

andrewkchan/yalm
3.6

Yet Another Language Model: LLM inference in C++/CUDA, no libraries except for I/O

Inference Engines
59664C++11mo ago

ICLR2025-Papers-with-Code

yinizhilian/ICLR2025-Papers-with-Code
3.3

历年ICLR论文和开源项目合集,包含ICLR2021、ICLR2022、ICLR2023、ICLR2024、ICLR2025.

Fine-tuning Tools
590341y ago

Athena-Public

winstonkoh87/Athena-Public
6.1

The Linux OS for AI Agents — Persistent memory, autonomy, and time-awareness for any LLM. Own the state. Rent the intelligence.

LLM Frameworks
58377PythonMITtoday

LLM-FineTuning-Large-Language-Models

rohan-paul/LLM-FineTuning-Large-Language-Models
3.6

LLM (Large Language Model) FineTuning

Fine-tuning Tools
578136Jupyter Notebook1y ago

KuiperLLama

zjhellofss/KuiperLLama
3.9

校招、秋招、春招、实习好项目,带你从零动手实现支持LLama2/3和Qwen2.5的大模型推理框架。

Inference Engines
572143C++10mo ago

langtest

Pacific-AI-Corp/langtest
6.3

Deliver safe & effective language models

Evaluation & Testing
55952PythonApache-2.08d ago

langtest

PacificAI/langtest
6.3

Deliver safe & effective language models

Evaluation & Testing
55952PythonApache-2.08d ago

pinferencia

underneathall/pinferencia
4.7

Python + Inference - Model Deployment library in Python. Simplest model inference server ever.

Model Serving
54383PythonApache-2.03y ago

fakecloud

faiscadev/fakecloud
6.0

Free, open-source AWS emulator. LocalStack alternative: 26 services, 1,924 operations, 100% conformance. No account, no auth token, no paid tier.

Evaluation & Testing
53540RustAGPL-3.04d ago

tessera

zengxiao-he/tessera
4.3

From teacher to tiles — a from-scratch LLM distillation & serving engine: custom Triton/CUDA kernels, FSDP distillation, paged-KV continuous batching, speculative decoding, a Rust gateway, a JAX oracle, and interpretability tooling.

Inference Engines
5309PythonNOASSERTION2mo ago

continuous-eval

relari-ai/continuous-eval
5.6

Data-Driven Evaluation for LLM-Powered Applications

Evaluation & Testing
51638PythonApache-2.017d ago

ome

ome-projects/ome
6.1

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

Model Serving
49692GoApache-2.0today

LLM-VM

anarchy-ai/LLM-VM
4.7

irresponsible innovation. Try now at https://chat.dev/

Fine-tuning Tools
490139PythonMIT2y ago

agency

operand/agency
4.8

A fast and minimal framework for building agentic systems

Agent Frameworks
48928PythonMIT2mo ago

awsome-distributed-training

awslabs/awsome-distributed-training
5.6

Collection of best practices, reference architectures, model training examples and utilities to train large models on AWS.

Fine-tuning Tools
472206ShellMIT-0today

Finetune_LLMs

mallorbc/Finetune_LLMs
3.8

Repo for fine-tuning Casual LLMs

Fine-tuning Tools
46786PythonAGPL-3.02y ago

JetStream

AI-Hypercomputer/JetStream
4.8

JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs welcome).

Model Serving
45567PythonApache-2.07mo ago

agentsilex

howl-anderson/agentsilex
4.7

A transparent, minimal, and hackable agent framework. ~300 lines of readable code. Full control, no magic.

Agent Frameworks
45446PythonMIT7mo ago

Aquila2

FlagAI-Open/Aquila2
3.6

The official repo of Aquila2 series proposed by BAAI, including pretrained & chat large language models.

Fine-tuning Tools
44732Python1y ago

xFasterTransformer

intel/xFasterTransformer
4.2

xFasterTransformer — open-source AI/LLM project.

Model Serving
43576C++Apache-2.011mo ago

gpu-rest-engine

NVIDIA/gpu-rest-engine
3.7

A REST API for Caffe using Docker and Go

Model Serving
42295C++BSD-3-Clause8y ago

InternEvo

InternLM/InternEvo
4.7

InternEvo is an open-sourced lightweight training framework aims to support model pre-training without the need for extensive dependencies.

Fine-tuning Tools
42167PythonApache-2.01y ago

Awesome-LLM-Prompt-Optimization

jxzhangjhu/Awesome-LLM-Prompt-Optimization
4.1

Awesome-LLM-Prompt-Optimization: a curated list of advanced prompt optimization and tuning methods in Large Language Models

Prompt Engineering
416242mo ago

awesome-azure-openai-llm

kimtth/awesome-azure-openai-llm
4.6

A curated collection of resources for 🌌 Azure OpenAI, 🦙 LLMs (RAG, Agents).

Agent Frameworks
40959Python10d ago

LightRFT

opendilab/LightRFT
5.2

LightRFT: Light, Efficient, Omni-modal & Reward-model Driven Reinforcement Fine-Tuning Framework

Fine-tuning Tools
40411PythonApache-2.011d ago

tiger

tigerlab-ai/tiger
4.3

Open Source LLM toolkit to build trustworthy LLM applications. TigerArmor (AI safety), TigerRAG (embedding, RAG), TigerTune (fine-tuning)

Fine-tuning Tools
40427Jupyter NotebookApache-2.02y ago

stable-diffusion-deploy

Lightning-Universe/stable-diffusion-deploy
4.6

Learn to serve Stable Diffusion models on cloud infrastructure at scale. This Lightning App shows load-balancing, orchestrating, pre-provisioning, dynamic batching, GPU-inference, micro-services working together via the Lightning Apps framework.

Model Serving
39138PythonApache-2.02y ago

rhesis

rhesis-ai/rhesis
5.6

The testing platform for AI teams. Bring engineers, PMs, and domain experts together to generate tests, simulate (adversarial) conversations, and trace every failure to its root cause.

Evaluation & Testing
39032PythonNOASSERTIONtoday

openclaw-optimization-guide

OnlyTerp/openclaw-optimization-guide
4.9

Make your OpenClaw AI agent faster, smarter, and cheaper. Speed optimization, memory architecture, context management, model selection, and one-shot development guide.

Prompt Engineering
38443JavaScriptMIT1mo ago

TensorSharp

zhongkaifu/TensorSharp
5.8

A native .NET LLM inference engine for GGUF models. TensorSharp provides a console application, a web-based chatbot interface, and Ollama/OpenAI-compatible HTTP APIs for programmatic access. It supports Windows/MacOS/Linux with full GPU capability

Inference Engines
37738C#BSD-3-Clausetoday

APOLLO

zhuhanqing/APOLLO
4.1

APOLLO: SGD-like Memory, AdamW-level Performance; MLSys'25 Oustanding Paper Honorable Mention

Fine-tuning Tools
36519PythonNOASSERTION9mo ago

flutter-skill

ai-dashboad/flutter-skill
5.5

AI-powered E2E testing for 10 platforms. 253 MCP tools. Zero config. Works with Claude, Cursor, Windsurf, Copilot. Test Flutter, React Native, iOS, Android, Web, Electron, Tauri, KMP, .NET MAUI — all from natural language.

Evaluation & Testing
36052DartMIT7d ago

Dulus

KevRojo/Dulus
5.3

Open-source autonomous AI agent — runs on Claude-web, Gemini-web, Kimi-web, Deepseek-web and more for free and Every paid model via liteLLM. No API key required.

Agent Frameworks
35820PythonGPL-3.01d ago

llm-leaderboard

JonathanChavezTamales/llm-leaderboard
4.6

A comprehensive set of LLM benchmark scores and provider prices. (deprecated, read more in README)

Evaluation & Testing
35740JavaScriptNOASSERTION10mo ago

alphora

opencmit/alphora
5.2

A Production-Ready Framework for Building Composable AI Agents

Agent Frameworks
34833PythonApache-2.01mo ago

Awesome-MLSys-Blogger

MLSys-Learner-Resources/Awesome-MLSys-Blogger
3.6

The repository has collected a batch of noteworthy MLSys bloggers (Algorithms/Systems)

Fine-tuning Tools
3449HTML1y ago

palico-ai

palico-ai/palico-ai
4.5

Build, Improve Performance, and Productionize your LLM Application with an Integrated Framework

Evaluation & Testing
34331TypeScriptMIT1y ago

vibe-log-cli

vibe-log/vibe-log-cli
5.0

A CLI tool for logging and analyzing Claude Code and Cursor ai-driven coding session.

Prompt Engineering
34021TypeScriptMIT4mo ago

ReaLHF

openpsi-project/ReaLHF
3.7

Super-Efficient RLHF Training of LLMs with Parameter Reallocation

Fine-tuning Tools
33622PythonApache-2.01y ago

swiftLLM

interestingLSY/swiftLLM
3.8

A tiny yet powerful LLM inference system tailored for researching purpose. vLLM-equivalent performance with only 2k lines of code (2% of vLLM).

Inference Engines
33337PythonApache-2.01y ago

memra

avifenesh/memra
5.9

Rust + CUDA inference engine for NVIDIA RTX PRO 6000 Blackwell and RTX 5090. Serves safetensors and GGUF over an OpenAI-compatible API, with per-device tuned defaults and speculative decode gated byte-identical to plain decode. Hosted instance: inference.tiyuvta.ai

Inference Engines
32738RustMITtoday

llms-tools

PetroIvaniuk/llms-tools
4.8

A list of LLMs Tools & Projects

Evaluation & Testing
32750Apache-2.029d ago

pmetal

Epistates/pmetal
4.8

PMetal: high-performance Apple Silicon framework for local LLM inference, LoRA/QLoRA fine-tuning, serving, quantization, and MLX/Metal acceleration.

Model Serving
31026RustNOASSERTION2mo ago