Ollama
Local LLM RunnersRun large language models locally with a simple CLI and API
Replaces ChatGPT, OpenAI API
The 9 best Lightweight local llm runners you can self-host, ranked by community traction.
Run large language models locally with a simple CLI and API
Replaces ChatGPT, OpenAI API
High-performance LLM inference in plain C/C++
Replaces OpenAI API
Distribute and run LLMs with a single executable file
Replaces OpenAI API, ChatGPT
Single-file local LLM runner for text and storytelling
Replaces ChatGPT
Browser extension to use local AI models on the web
Replaces ChatGPT
Self-hosted LLaMA chat UI with no API keys needed
Replaces ChatGPT
Fast inference server for text embedding models
Replaces OpenAI Embeddings API
Performance-focused fork of llama.cpp with new quant types
Replaces OpenAI API
Run LLM inference directly in the browser with WebAssembly
Replaces OpenAI API
No apps match these filters.
Every option here is open-source and self-hostable, tagged Lightweight within the local llm runners category. Compare them on the individual app pages for setup difficulty and resource requirements.