Running language and image models on your own laptop, phone or server instead of a cloud API. Named by Georgi Gerganov: llama.cpp: LLaMA inference in plain C/C++ (Mar 10, 2023). Also called Local LLMs, On-device AI, Self-hosted AI, Running models locally.
Index 0 to 100, monthly
Signals
Signal
Value
Reads as
Hacker News stories, last 12 months813 stories, +185% on the 12 months before
813 stories, +185% on the 12 months before
Adoption
arXiv papers, last 12 months237 papers, +81% on the 12 months before
nomic-ai/gpt4allGPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use. · Mar 27, 2023 3y3 years ago
Mar 27, 20233y3 years ago
77,384
oobabooga/textgenOpen-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private. · Dec 21, 2022 3y3 years ago
Dec 21, 20223y3 years ago
47,706
khoj-ai/khojYour AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Tur · Aug 16, 2021 5y5 years ago
Aug 16, 20215y5 years ago
37,493
intel/ipex-llmAccelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, et · Aug 29, 2016 10y10 years ago
Aug 29, 201610y10 years ago
8,853
Andyyyy64/whichllmFind the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One · Mar 4, 2026 6mo6 months ago
Mar 4, 20266mo6 months ago
6,682
pguso/ai-agents-from-scratchDemystify AI agents by building them yourself. Local LLMs, no black boxes, real understanding of function calling, memory, and ReAct pattern · Oct 23, 2025 11mo11 months ago
Oct 23, 202511mo11 months ago
4,809
Related movements
Movement
Stage
Small language modelsCompact models, a few billion parameters or less, trained to be good enough and cheap to run anywhere.
Peak
Open-weight modelsModels whose trained weights anyone can download, run and fine-tune: Llama, Qwen, DeepSeek, Mistral and more.
Rising
Inference chipsAccelerators built to serve models cheaply and fast rather than to train them: TPUs, LPUs, wafer-scale and custom silicon.
Peak
History
Date
What changed
Sep 24, 2026
Tracking started: emerged March 2023, peak, 4 other names recorded
Sources: Curve: Hacker News story titles, Wikipedia pageviews and arXiv papers by month, refreshed Sep 24, 2026. Related entities from the fru.dev sites' public APIs, matched by name. How stages work.
Trending terms by email
Monday mornings: the week's top trending terms in data, tech and AI.