The AI movement's repository database

The best AI repositories

224 living projects — from agent frameworks to prompt-injection defense. Each comes with our own note: why it matters, who it is for and how to start in three steps.

Stars, forks and activity refresh daily via the GitHub API · updated October 6, 2026

repositories
224
categories
11
stars in total
11M
new stars in 30 days
+235K

Last 30 days

Growing fastest

The most new stars over the last 30 days.

Full growth ranking
Growing fastestRookie of the month

A skills set and methodology for coding agents: clarify the task, plan, write tests, then code. The agent works with more discipline.

296Kstars+15K in 30 dShell

firecrawl

firecrawl

Growing fastest

A service that turns any website into clean Markdown or structured data for models. It can search, scrape, crawl a whole site and even click through a page.

189Kstars+12K in 30 dTypeScript

microsoft

markitdown

Growing fastest

A small Microsoft utility that converts PDF, Word, PowerPoint, Excel and other files into Markdown that is easy to hand to a language model.

189Kstars+11K in 30 dPython

anomalyco

opencode

Growing fastest

An open coding agent for the terminal with a desktop app. It is not tied to one provider: plug in whichever models you have.

212Kstars+7.6K in 30 dTypeScript

github

spec-kit

Growing fastest

A toolkit for spec-driven development: principles and a spec first, then a plan and tasks, and only then code. Works with several coding agents.

140Kstars+7.1K in 30 dPython

openai

codex

Growing fastest

OpenAI's lightweight coding agent that runs in the terminal and works on your project's code. Written in Rust.

128Kstars+6.8K in 30 dRust

Categories

Agent frameworksFrameworks and orchestration: what agents and multi-agent systems are built from.23 repositoriesCoding agents & CLIAgents that write code in the terminal and editor, and the methods for working with them.21 repositoriesMCPThe Model Context Protocol spec, official SDKs and the best servers.20 repositoriesOpen models & inferenceOpen weights, local runs, fast inference and fine-tuning.38 repositoriesRAG & vector DBsSearch over your own data: document parsing, crawlers, vector databases, knowledge graphs.19 repositoriesEvals & observabilityHow to measure a model or agent and see what happens inside it.17 repositoriesPrompting & contextPrompts, structured output, skills and context engineering.16 repositoriesVoice, video, multimodalSpeech recognition and synthesis, voice agents, images and video.19 repositoriesAutomationNo-code workflows, browser and computer-use agents.15 repositoriesAwesome lists & learningCourses, books and awesome lists worth starting with.21 repositoriesSecurityPrompt injection, guardrails, red teaming and agent scanners.15 repositoriesMissing your favorite?Suggest a repository — we will review it and add it if it earns a spot.Suggest a repo →

Catalog

19 repositories

The classic browser UI for Stable Diffusion: text-to-image, inpainting, upscaling and thousands of community extensions.

165Kstars+664 in 30 d33KPythonCommit 7 months ago

ggml-org

llama.cpp

A C/C++ engine that runs language models on ordinary hardware: laptops, phones, GPU-less servers. Much of local AI, Ollama included, is built on it.

130Kstars+3.6K in 30 d24KC++Commit today

pytorch

pytorch

The foundation under most modern neural networks: GPU-accelerated tensors and automatic differentiation, wrapped in ordinary Python. Transformers, Llama and nearly every open model are written on it.

104Kstars+1.1K in 30 d31KPythonCommit 1 day ago

vllm-project

vllm

An engine for fast model serving on GPUs: frugal with memory and handles many requests at once. The default pick when a model has to run as a service.

93Kstars+2.4K in 30 d23KPythonCommit today

unslothai

unsloth

A tool for running and fine-tuning models while saving memory: there is a desktop app, a web UI and a library. Fine-tuning works on a modest GPU.

77Kstars+1.7K in 30 d7.1KPythonCommit today

karpathy

nanochat

Rookie of the month

The whole path from zero to your own ChatGPT-like chat in one repo: tokenizer, pretraining, fine-tuning and UI. Designed to train for roughly a hundred dollars.

58Kstars+753 in 30 d8.2KPythonCommit 28 days ago

mudler

LocalAI

A local replacement for cloud APIs: an OpenAI-compatible server that runs text, voice and image models on any hardware, no GPU required.

49Kstars+575 in 30 d4.5KGoCommit 1 day ago

facebookresearch

faiss

Meta's classic library for fast similarity search and clustering of vectors. It is not a database but an index engine that many other systems are built on.

41Kstars+238 in 30 d4.6KC++Commit 3 days ago

sgl-project

sglang

A fast engine for serving language and multimodal models. A vLLM rival, especially strong on large models like DeepSeek.

37Kstars+1.4K in 30 d9.3KPythonCommit 1 day ago

karpathy

llm.c

GPT-2 training in plain C and CUDA, no PyTorch. Shows what happens under the hood while a model learns.

31Kstars+211 in 30 d3.8KCudaCommit 1 year ago

black-forest-labs

flux

The official inference code for FLUX.1 models: text-to-image generation and editing, including Kontext mode where you edit an image with words.

26Kstars+113 in 30 d1.9KPythonCommit 1 year ago

Whisper on the CTranslate2 engine: same output, noticeably faster and lighter on memory, including 8-bit mode.

26Kstars+493 in 30 d2.1KPythonCommit 1 day ago

A fast, memory-efficient implementation of attention that large-model training and inference on GPUs lean on. The result is exact, with no approximation.

25Kstars+269 in 30 d3.1KPythonCommit 1 day ago

state-spaces

mamba

An alternative to the transformer: a state space model whose runtime grows linearly with text length. A good fit for long sequences.

19Kstars+103 in 30 d1.8KPythonCommit 3 months ago

Wan-Video

Wan2.1

Open video-generation models: turn text or an image into a clip. The 1.3B version fits in roughly 8 GB of VRAM.

17Kstars+194 in 30 d3.7KPythonCommit 7 months ago

ggml-org

ggml

A small tensor library in plain C/C++ with no dependencies. It underpins llama.cpp and whisper.cpp and runs quantized models on CPU, GPU and in the browser.

15Kstars+163 in 30 d1.9KC++Commit 1 day ago

The standard toolkit for running models through hundreds of academic benchmarks. Many open model leaderboards are built on it.

14Kstars+242 in 30 d3.6KPythonCommit 22 days ago

bitsandbytes-foundation

bitsandbytes

Squeezes models down to 8 and 4 bits so they fit on an ordinary GPU. QLoRA and running large models in half the memory both rely on it.

8.5Kstars+57 in 30 d938PythonCommit 28 days ago

EleutherAI

gpt-neox

EleutherAI's library for training large language models from scratch across many GPUs: ZeRO and 3D parallelism, launching via Slurm and MPI. Its own README says to use it only for models with billions of parameters.

7.5Kstars+13 in 27 d1.1KPythonCommit 31 days ago

Did we miss something?

Suggest a repository

Send a GitHub link and a few words on why it belongs here. Every suggestion is reviewed by hand — not everything gets in.