OpenClaw
Self-Hosted AIThe AI that actually does things
Run large language models, chat UIs, and machine-learning tools locally with self-hosted AI software.
137 self-hosted apps · 328 comparisons
The AI that actually does things
The AI agent that grows with you
The open source AI coding agent
Replaces Claude Code, Cursor
State-of-the-art machine learning model library
Replaces OpenAI API
Open-source platform for building production LLM apps
Replaces OpenAI Assistants, Vertex AI Agent Builder
Visual framework for building AI agents and RAG pipelines
Replaces Vertex AI Agent Builder
RAG engine with deep document understanding
Replaces NotebookLM
Practical multilingual OCR toolkit
Replaces Google Cloud Vision, AWS Textract
Autonomous AI software engineering agent
Replaces Devin, GitHub Copilot Workspace
Convert screenshots and designs into working code
Replaces Vercel v0
Fast, memory-efficient LLM finetuning
Replaces OpenAI Fine-tuning
LLM toolkit for research, papers, and code analysis
Replaces ChatGPT
Autonomous coding agent for your editor
Replaces GitHub Copilot, Cursor
Natural-language interface that runs code on your machine
Replaces ChatGPT Code Interpreter
Parse documents into AI-ready structured formats
Replaces Adobe Acrobat
Unified OpenAI-compatible API gateway for many models
Replaces OpenAI API
All-in-one desktop and server RAG application
Replaces ChatGPT, NotebookLM
Self-hosted memory layer for AI agents and assistants
Replaces OpenAI Memory
Few-shot voice cloning and text-to-speech
Replaces ElevenLabs
Ask questions about your documents fully offline
Replaces ChatGPT
Drag-and-drop builder for LLM apps and AI agents
Replaces Vertex AI Agent Builder
Generate entire codebases from a natural language prompt
Replaces GitHub Copilot Workspace
Fast C++ port of OpenAI Whisper speech recognition
Replaces Otter.ai, Google Speech-to-Text
Free AI image upscaler for the desktop
Replaces Topaz Gigapixel AI
AI pair programming in your terminal
Replaces GitHub Copilot, Cursor
Gradio web UI for running and tuning local text models
Replaces ChatGPT
Deep learning toolkit for local text-to-speech
Replaces ElevenLabs
Open-source framework for augmenting humans with AI
Replaces ChatGPT
Lightweight framework for building multi-agent systems
Replaces LangChain
Convert PDFs and documents to Markdown with AI
Replaces Adobe Acrobat
Simple and fast retrieval-augmented generation with graphs
Replaces OpenAI Assistants
Opinionated RAG framework for your second brain
Replaces NotebookLM
Self-hosted AI notebook for documents and audio overviews
Replaces Google NotebookLM
Transformer-based text-to-audio generation model
Replaces ElevenLabs
Self-hosted voice conversion and cloning interface
Replaces ElevenLabs voice cloning
Framework for programming, not prompting, language models
Replaces LangChain
Instant voice cloning with flexible style control
Replaces ElevenLabs
Fast serving framework for LLMs and vision-language models
Replaces OpenAI API
Configure and deploy autonomous AI agents in the browser
Replaces AutoGPT
Graph-based retrieval-augmented generation system
Replaces OpenAI Assistants
Self-hosted tool to test and refine LLM prompts
Replaces PromptPerfect
Open-source observability and analytics for LLM apps
Replaces LangSmith
Self-hosted AI coding assistant alternative to Copilot
Replaces GitHub Copilot
Customizable frontend for AI character roleplay
Replaces Character.AI
AI developer that writes apps with human oversight
Replaces Devin
AI assistant connected to your company's knowledge
Replaces Glean, ChatGPT Enterprise
Memory and knowledge graph engine for AI agents
Replaces Mem0, Pinecone
Ultra-lightweight personal AI assistant
Industry-leading face manipulation and swapping platform
Replaces DeepFaceLab
Autonomous agent for comprehensive online research reports
Replaces Perplexity, ChatGPT
Open-source embedding database for AI applications
Replaces Pinecone
Framework for building production-ready LLM applications
Replaces LangChain
Clean RAG UI for chatting with your documents
Replaces NotebookLM
Reimplementation of Whisper for fast local transcription
Replaces Google Speech-to-Text
Test, evaluate, and red-team LLM applications
Replaces LangSmith
Build stateful LLM agents with long-term memory
Replaces OpenAI Assistants
Remove image backgrounds automatically with AI
Replaces remove.bg
Customizable AI dev team inside VS Code
Replaces GitHub Copilot, Cursor
Fast Whisper transcription with word-level timestamps and diarization
Replaces Otter.ai, Rev
Chat with your SQL database using RAG-trained models
Replaces ThoughtSpot
Generative audio and music models from Meta
Replaces Suno, ElevenLabs
Knowledge base question-answering system built on LLMs
Replaces Intercom Fin
Describe UI with text and see it rendered live
Replaces Vercel v0
Chat with your documents fully offline
Replaces ChatGPT
Constrained generation language for controlling LLMs
Replaces OpenAI API
Multilingual document OCR and layout analysis toolkit
Replaces Google Cloud Vision, AWS Textract
Framework for building contextual text and voice assistants
Replaces Dialogflow, Amazon Lex
Language-model agent that resolves GitHub issues
Replaces Devin
Personal organic agentic AI framework that grows with use
Replaces AutoGPT, Devin
Unit testing framework for LLM outputs
Replaces LangSmith, Braintrust
Framework for building controlled conversational AI agents
Replaces Voiceflow
Open-source GenBI agent for talking to your database
Replaces ThoughtSpot
Open-source framework for autonomous AI agents
Replaces AutoGPT
Open-source personal AI assistant you can self-host
Replaces Google Assistant, Alexa
Free and Open Source machine translation API
Replaces Google Translate, DeepL
Query multiple AI chatbots side by side at once
Replaces Poe
Customizable AI research agent connected to your knowledge
Replaces NotebookLM, Perplexity
Structured text generation for language models
Replaces OpenAI Structured Outputs
Evaluation framework for retrieval-augmented generation
Replaces LangSmith
Framework for enterprise RAG with small language models
Replaces OpenAI Assistants
Unified framework to benchmark language models on many tasks
Replaces OpenAI Evals
Recipes for pretraining and finetuning open LLMs
Replaces OpenAI Fine-tuning
All-in-one embeddings database for semantic search and RAG
Replaces Pinecone
Open LLM application platform for enterprise workflows
Replaces Dify
Private chat and document Q&A with local LLMs
Replaces ChatGPT
Open-source AI observability and evaluation platform
Replaces LangSmith, Helicone
Embedded vector database for AI applications
Replaces Pinecone
Build and evaluate LLM application flows
Replaces LangSmith, Vertex AI
High-performance inference serving for any model framework
Replaces Amazon SageMaker
Self-hosted offline ChatGPT-like chatbot powered by Llama
Replaces ChatGPT
All-in-one LLM CLI tool with chat-REPL and API server
Replaces ChatGPT
High-performance RLHF framework for training LLMs
Replaces OpenAI Fine-tuning
Distributed inference framework for LLMs and embeddings
Replaces OpenAI API, Hugging Face Inference Endpoints
AI-powered deep research tool with multi-source search (arXiv, PubMed,
Compact, high-quality open text-to-speech model
Replaces ElevenLabs, OpenAI TTS
Framework for building and serving AI model APIs
Replaces Amazon SageMaker, Vertex AI
Composable API server for building generative AI applications
Replaces OpenAI API
Production-ready RAG engine with a RESTful API
Replaces OpenAI Assistants
AI assistant that turns issues into pull requests
Replaces GitHub Copilot Workspace
Open-source RAG chatbot powered by Weaviate
Replaces NotebookLM
High-quality multilingual text-to-speech library
Replaces ElevenLabs, Amazon Polly
Open-source observability for LLM applications
Replaces LangSmith, Datadog LLM Observability
Advanced all-in-one generative image and video studio
Replaces Midjourney
Run and fine-tune LLMs locally on Apple Silicon with MLX
Replaces OpenAI API
Manage GPU clusters for running AI models
Replaces OpenAI API
Workspace for training and evaluating local LLMs
Replaces OpenAI Playground
Hub for many AI models with a personality system
Replaces ChatGPT
Open-source LLMOps platform for prompt and app management
Replaces LangSmith
Visual programming environment for building AI agents
Replaces LangChain
Fast inference library for quantized LLMs on consumer GPUs
Replaces OpenAI API
Modular framework for production-ready RAG systems
Replaces LangChain
Simple, hackable GraphRAG implementation in Python
Replaces Microsoft GraphRAG
OpenAI-compatible speech-to-text and text-to-speech server
Replaces OpenAI Audio API, ElevenLabs
Framework for synthetic data generation and AI feedback
Replaces Scale AI
Self-hosted speech-to-text API powered by OpenAI Whisper
Replaces Google Speech-to-Text, Otter.ai
State-of-the-art music source separation
Replaces LALAL.AI
Open-source privacy-focused voice assistant hardware platform
Replaces Amazon Alexa, Google Assistant
Framework for building AI assistants with memory and plugins
Replaces OpenAI Assistants
Run AI models using OCI containers
Replaces Ollama
Self-hosted face swapping for images and videos
Replaces Reface
Memory-efficient inference library for quantized Llama models
Replaces OpenAI API
DSPy-inspired framework for building typed LLM programs
Replaces DSPy
Quickly deploy your own ChatGPT-style web service
Replaces ChatGPT
AI agent for teams, communities, and multi-user environments
On-premise generative AI platform for teams
Replaces ChatGPT Enterprise
Self-hosted autonomous AI agent platform
Replaces AutoGPT
Polished web UI for ChatGPT, DALL-E and other AI models
Replaces ChatGPT
Private AI chat and document analysis you fully control
Replaces ChatGPT
A ChatGPT web client that supports multiple users, multiple languages,
Crowdsourced distributed cluster for AI generation
Replaces Midjourney
Toolkit for tracking and evaluating LLM applications
Replaces LangSmith
Chat with your documents using local AI
Replaces NotebookLM
Self-hosted virtual character chat with voice and 3D models
Replaces Character.AI
OpenAI-compatible text-to-speech server using local models
Replaces OpenAI API, ElevenLabs
A document reader with Text-to-Speech capabilities
Speech-to-text service for the Wyoming voice protocol
Replaces Google Speech-to-Text
Framework for training instruction-tuned language models
Replaces OpenAI Fine-tuning
No apps match these filters.
Last reviewed Aug 26, 2026 · 454 words
The 138 tools filed under "self-hosted AI" are really four shelves in one aisle: engines that run model weights, chat frontends for talking to them, visual builders for assembling LLM apps, and agents that go off and do things on your behalf. Most bad installs in this category are shelf errors — someone wanted a private ChatGPT and ended up deploying a workflow orchestrator with a Postgres dependency. Figure out which shelf you are on and the shortlist mostly picks itself.
Three axes do the splitting. First, the layer: if you want to chat, you need a runner from Local LLM Runners plus a frontend from AI Chat UIs — nothing in this parent category is that. Second, where the intelligence actually lives: a surprising share of "self-hosted AI", including the biggest agents here, is orchestration around a paid API key, so your prompts still leave the building unless you deliberately point the tool at local models. Third, hardware: agent shells idle at 512 MB, app platforms like Dify want 4 GB, and anything that loads model weights itself starts at 8 GB and climbs from there.
For teams building LLM features, Dify (Apache-2.0, 153,524 stars) bundles a visual workflow builder, RAG pipelines, and observability in one multi-container stack — heavier than a toy, but it replaces three or four separate tools. For a coding agent, OpenHands (MIT, 4 GB) writes code, runs commands, and browses the web inside a sandboxed Docker runtime, and works against local or remote models; budget for token costs if you use a hosted provider.
OpenClaw tops this category at 387,634 stars, and it is a genuinely useful assistant — inbox, calendar, flight check-ins, driven from WhatsApp or Telegram. It is also the clearest example of the category's asterisk: it requires a model API key, ships a headless browser with real reach into your accounts, and its own docs stress that security needs careful setup. The hottest projects here are agents orchestrating cloud models, not local inference. That is a fine trade, but make it on purpose, and give any agent the least access you can get away with.
Run a small model locally with a runner and a chat UI first — it is a weekend project and teaches you the RAM and latency realities everything else in this category inherits. Once that works, my default second install is Dify: broad enough to cover RAG, agents, and workflows, and boring enough (Apache-2.0, Docker Compose) to still be running next year.
328 head-to-head comparisons in this category.