Research desks that cite your files, agents that run your desktop, whole offline suites โ plus the plumbing to route it all
Private AI on your own machine โ offline, no subscription, nothing you type ever leaving the room. 90+ tools, sorted โ the rarest first:
ANARCHY โ a fully offline AI shell that runs your desktop: voice, RAG, Git, automation, permission tiers.
Qube โ research-grade private answers with real, inspectable citations, on its own engine.
Odysseus โ a whole private suite: chat, agents, email, notes, calendar โ all local.
lilbee โ an entire local AI stack (runtime + vector index + crawler) in one executable.
Akana โ a self-hosted server where every memory write needs your approval, behind an encrypted vault.
The stack โ private AI on your own hardware
๐งฑ The engine underneath โ pick one, the rest runs on it
- Ollama โ one command to pull + run a model; the backbone most tools sit on.
- LM Studio โ friendly desktop model lab plus a local API server.
- Jan โ offline ChatGPT-style desktop app, no terminal.
- Open WebUI โ the ChatGPT-style browser front-end for any of them.
- llama.cpp โ the raw engine underneath; runs on CPU when thereโs no GPU worth using.
๐ Chat with your own files, privately โ cited answers, no cloud
- Qube โ native desktop assistant, own GGUF engine, editable memory, inspectable citations.
- lilbee โ a whole local stack (runtime + vector index + crawler) in one executable.
- Chaty โ offline GGUF/MLX chat with RAG, voice, and deep-research.
- GGUF Loader โ offline models that act on a project folder, with cited RAG.
- Beacon โ Ollama + LanceDB + reranking, auditable answers from your files.
- Khoj โ a self-hostable โsecond brainโ over your docs, custom agents, local backends.
- AnythingLLM โ self-hostable RAG workspaces + agents over local providers.
- Gemma Genie โ CLI Q&A over PDFs, Office files, images and audio, on-device.
๐ค AI that acts on your machine โ agent desks, not chat boxes
- ANARCHY โ fully local shell: voice, RAG, Git, desktop automation, permission tiers.
- Feral โ hackable local agent desk with a sidecar runtime + deep-research.
- nanobot โ build a private assistant network: WebUI, tools, memory, MCP, connectors.
- Local Agent Studio โ wires local Ollama + ComfyUI with sandboxed tool execution.
- Locally Uncensored โ one installer: chat, coding, image/video, voice, RAG.
- Pengy โ one shell (GUI/CLI/web) routing several local model servers.
- DARIOS AI โ voice + local LLM + document memory + DevSecOps tools.
๐๏ธ A whole private workspace โ suites that replace the SaaS, not just the chat
- Odysseus โ chat, agents, research, docs, email, notes, calendar โ all local.
- Artha โ a local agent that makes DOCX/XLSX/PDF files; an offline office assistant.
- Akana โ self-hosted server with review-gated memory + an encrypted vault.
- Aster โ private workspace with bounded, persistent agents.
- Off Grid AI โ local studio: llama.cpp inference, OpenAI-compatible gateway, transcription, image gen.
๐ง Serve the whole household โ one local box, every device
- Local AI โ portable Ollama workspace, RAM-aware, LAN-served to phones + laptops.
- BlackCortex Lite โ one Windows box serves the family over the LAN, per-user history, PDF chat.
- rag_unplugged โ Dockerized stack (Ollama, ChromaDB, SearXNG, Telegram/WhatsApp), every block swappable.
- rag-chatbot โ a Raspberry-Pi RAG bot for a household or small office.
๐งโ๐ป Lean, air-gapped & creative โ small binaries, companions, multi-model
- Claudette โ single-binary coding agent; tested offline mode blocks outbound calls (air-gap-safe).
- Pern โ offline dev workspace with an embedded Llama server.
- Sullybase โ lean Flask chat with live hardware monitoring.
- PolyUI โ compare several local models side by side, offline.
- SweetrollLM ยท Aria โ offline character/companion chat, no cloud accounts.
- XandSuite โ an all-in-one offline playground: GGUF, RAG, multimodal tools.
๐ No machine to spare? No-install browser chats (no card, no signup)
- AIFreeForever โ many model-style chats + search + files + image gen, free text, no login.
- Free.ai ยท AskingTips ยท Tusk Central ยท PLAI โ multi-model hubs, no account.
- Duck.ai ยท Privatemode ยท notrack.ai โ privacy-first, anonymous, no tracking.
- BrowserLLM ยท OrangeBot ยท Private AI Chat โ the model runs inside your browser (WebGPU), nothing sent out.
- Free LLM Playground ยท GLM-AI โ no-signup testing / GLM-family chat.
๐ฏ Strong models, free in the browser โ when you want frontier quality now
- DeepSeek Chat ยท DeepSeek V4 Flash โ capable free chat; V4 Flash is no-signup.
- Mistral Le Chat โ serious European-hosted free chatbot.
- HuggingChat โ open-model chat, guest use.
- Perplexity โ web-grounded research answers, free tier.
- LMArena โ pit many live models against your task, free.
- Rewind ยท ComfyAI ยท Kumori โ daily-free / no-cap / shared-room chats.
๐ The plumbing โ run, route & manage your local models
- Route & serve: llama.cpp Control Deck ยท Local Model Router ยท SmarterRouter (semantic cache + failover) ยท lmswitch ยท GGUF Desktop (OpenAI-compatible bridge).
- Manage & launch: Ollama Model Manager ยท LlamaDeck ยท Local-LLM-Launcher-GUI ยท Local Model Manager.
- Fit your hardware: AI Model Compass ยท LLMFit ยท hfo ยท minimal-ai ยท ZN-LLM (auto-tunes to your machine).
- Memory + safety: Localmind ยท pmb ยท Memgentic (share memory across tools) ยท PrivateRedact (offline PII redaction).
Your machine is the subscription now.

!