Ollama
One-command local LLM runtime.
Run large language models on your own hardware.
Select for model-format support, context needs, and the amount of CPU/GPU offload you can tolerate—not only for the largest model the launcher can download.
One-command local LLM runtime.
Agentic coding in VS Code — reads, writes, runs, browses.
Self-hosted ChatGPT-style frontend for Ollama / OpenAI.
Desktop GUI for running local LLMs.
The C++ inference engine powering most local LLMs.
Open-source Copilot — VS Code & JetBrains, any model.
Beautifully designed chat UI with plugins and image generation.
Open-source ChatGPT desktop — runs models locally or via API.
Whisper, 4× faster, same accuracy. CTranslate2 backend.
Polished local-LLM client with split chats and knowledge stacks.
Terminal-native AI pair programmer with git awareness.
RAG-first local LLM workspace with workspaces and agents.
The "A1111 for LLMs" — multi-loader local chat UI.
Natural-language code execution on your machine.
Open-weight text-to-audio — 47-second sound effects and music.