Tool for running large language models locally on your machine with a simple CLI interface. Download and run Llama 3, Mistral, Gemma, Phi, Code Llama, and dozens of other open-source models with a single command. Features model management, GPU acceleration (NVIDIA/AMD/Apple Silicon), OpenAI-compatible API server, Modelfile for customization, and multi-model switching. Ideal for offline AI development, privacy-sensitive use cases, and local testing. 120K+ GitHub stars.
Alternatives to AnythingLLM
7 editor-selected alternatives · AnythingLLM overview →
source: tools.alternatives · stored order · active records only; review scores are annotations and never change membership or order
A directional evidence panel appears only when the substitute rationale, trade-offs, sources, and verification date have been recorded. Older selections without that panel remain visible but are unclassified under the new evidence contract.
Extensible, self-hosted AI platform with 290M+ Docker pulls and 124K+ GitHub stars. Supports Ollama, OpenAI-compatible APIs, and any Chat Completions backend. Features built-in RAG, multi-user RBAC, voice/video calls, Python function workspace, model builder, and web browsing. Runs entirely offline with enterprise features including SSO and audit logging.
Jan is an open-source offline-first AI assistant with 25K+ GitHub stars running LLMs locally without sending data externally. Features a ChatGPT-like interface with one-click model downloads from Hugging Face, conversation management, customizable prompts, and an OpenAI-compatible local API server. Supports GGUF models via llama.cpp with GPU acceleration on NVIDIA and Apple Silicon. Built with Electron for macOS, Windows, and Linux with full data privacy.
LobeChat is a source-available AI chat and agent workspace for OpenAI, Claude, Gemini, Ollama, DeepSeek, and Qwen. It includes RAG, 10,000+ MCP-compatible plugins, Agent Groups, TTS/STT, Vercel/Docker self-hosting, and 79K+ GitHub stars.
Khoj is an open-source personal AI app that serves as a self-hostable second brain. It connects to your documents — PDFs, Markdown, Notion, Word — and uses RAG to answer questions grounded in your knowledge base. Supports any local or cloud LLM including Llama, Claude, GPT, and Gemini. Features custom agents, scheduled automations, deep research mode, semantic search, and Obsidian, Emacs, and WhatsApp integrations. Over 33,000 GitHub stars, YC-backed.
Onyx is an open-core, self-hostable AI knowledge platform for enterprise search, RAG chat, deep research, custom agents, and workplace connectors. It connects to 40+ apps, supports permission-aware retrieval, and offers Cloud, Docker/Kubernetes, and enterprise deployment paths for teams that need controlled internal AI search.
QMD is an on-device search engine built by Tobi Lütke (Shopify CEO) that indexes markdown notes, meeting transcripts, and documentation locally. It combines BM25 full-text search, vector semantic search, and LLM-powered re-ranking into a single hybrid pipeline. Ships with a built-in MCP server for seamless integration with Claude Code, Cursor, and other AI editors. All processing happens on your machine via node-llama-cpp with GGUF models — zero cloud dependency.
Open-source AnythingLLM alternatives
Ollama, Jan, LobeChat, Khoj, Onyx, QMD — see all open-source developer tools.
Free AnythingLLM alternatives
Onyx offer a free plan or free tier.
More Self-Hosted Platforms tools
same category, not editor-selected alternatives — see how AnythingLLM compares →
AnythingLLM head-to-head
FAQ
Which AnythingLLM alternative is listed first?
Ollama is first in the editor-selected list of 7 AnythingLLM alternatives and carries an editorial review score of 88/100. The stored order is editorial; review scores do not determine membership or position.
Are there open-source AnythingLLM alternatives?
Yes — Ollama, Jan, LobeChat, and more are open source.
Are there free AnythingLLM alternatives?
Yes — Onyx offer a free plan or free tier.