💻 Open source

Run Models Locally

Everything you need to run a capable model on your own hardware, offline and free.

12repositories
76kstars combined
LL

llamafile

mozilla-ai

Distribute and run LLMs with a single file.

★ 25k · C++cross-platformggufllama-cpp
updated today View on GitHub →
LO

local-deep-research

LearningCircuit

~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.

★ 9k · Python · MITacademiaanthropicarxiv
updated today View on GitHub →
OP

open-multi-agent

open-multi-agent

TypeScript AI agent orchestration framework with dynamic workflows. Describe the goal, not the graph: a coordinator plans the task DAG at runtime and runs it on any LLM (Claude, ChatGPT, Gemini, DeepSeek, or local models).

★ 7k · TypeScript · MITagent-frameworkagent-orchestrationagentic-ai
updated today View on GitHub →
WH

whichllm

Andyyyy64

Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.

★ 6k · Python · MITaiapple-siliconbenchmarks
updated 14 days ago View on GitHub →
DO

dograh

dograh-hq

Open source voice AI platform. Self-hosted alternative to Vapi and Retell. On Prem, BYOK across Speech to Speech or LLM/STT/TTS, with a visual workflow builder, MCP native and telephony support.

★ 5k · Python · BSD-2-Clauseai-callingasterisk-ariconversational-ai
updated today View on GitHub →
OP

openmed

maziyarpanahi

Local-first healthcare AI: clinical NER & HIPAA PII de-identification that runs 100% on-device. 1,000+ medical models, 12 languages, Apple MLX + Python, no cloud, no patient data leaving your network. Apache-2.0

★ 5k · Python · Apache-2.0clinical-nlphealthcarehipaa
updated today View on GitHub →
LA

langroid

langroid

Harness LLMs with Multi-Agent Programming

★ 4k · Python · MITagentsaichatgpt
updated 10 days ago View on GitHub →
SU

surf

deta

Personal AI Notebooks. Organize files & webpages and generate notes from them. Open source, local & open data, open model choice (incl. local).

★ 3k · TypeScript · Apache-2.0claudedeepseekgemma
updated 7 days ago View on GitHub →
RA

Rapid-MLX

raullenchai

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

★ 3k · Python · Apache-2.0apple-siliconclaude-codecursor
updated today View on GitHub →
AL

algernon

xyproto

Small self-contained pure-Go web server with Lua, Teal, Markdown, Ollama, HTTP/2, QUIC, Redis, TypeScript, SQLite and PostgreSQL support ++

★ 3k · JavaScript · BSD-3-Clausealgernonbuild-lesscross-platform
updated 4 days ago View on GitHub →
CL

claude-code-local

nicedreamzapp

Run Claude Code 100% on-device with local AI on Apple Silicon. MLX-native Anthropic-API server, 65 tok/s Qwen 3.5 122B, Llama 3.3 70B, Gemma 4 31B. Private, offline, airgap-ready. Built for NDA / legal / healthcare workflows.

★ 3k · Python · MITabliteratedai-privacyairgap
updated 3 days ago View on GitHub →

A curated list of awesome platforms, tools, practices and resources that helps run LLMs locally

★ 2k · MITaiawesomeawesome-list
updated today View on GitHub →

Where this comes from: GitHub search: topic:local-llm stars:>1000, sorted by stars. Last refreshed 22 July 2026. Nothing on this page is a paid placement.