LocalAI
★ 49.5kA self-hosted alternative to Chatgpt
MIT LicenseOpen Source — No PaywallSingle Container — single container — minimal hosting footprintDocker · Binary · SourceNo external DB17+ active committers (12mo)
About
LocalAI is an open-source AI engine designed for self-hosted use. It lets users run a broad range of model types locally, including text, vision, speech, image, and video workloads, while keeping...
Community Ratings
No ratings yetNew this week
Read recap →Weekly Recap — Oct 2, 2026 – Oct 9, 2026Oct 9, 2026
Features
- OpenAI-compatible APIs
- Anthropic API compatibility
- ElevenLabs API compatibility
- LLMs, vision, voice, image, and video support
- On-demand backend loading
- Multi-user API key auth
- User quotas
- Role-based access
- Built-in AI agents
- Terminal agent
Details
- Last Updated
- Oct 10, 2026
- Created
- Mar 18, 2023
- Install Methods
- dockerbinarysource
- Requirements
- macOS (for the DMG quickstart)Docker or Podman for container-based setupGPU hardware is optional; CPU-only is supported
- Authentication
- local
- API
- REST
- Runtime / Stack
- DockerK8S
- Privacy & Independence
- No Cloud RequiredOffline Capable
- User Model
- Multi-User
- Deployment
Deployment: Docker Compose ✓Single Container — minimal hosting footprintsingle container
Run LocalAI with Docker Compose
Compose file from the LocalAI project repository.
1 service
no resource limits declared
needs persistent storage (5 volumes)
docker-compose.yaml · @ 61f4f67 · scanned Sep 28, 2026
services:
api:
# See https://localai.io/basics/getting_started/#container-images for
# a list of available container images (or build your own with the provided Dockerfile)
# Available images with CUDA, ROCm, SYCL
# Image list (quay.io): https://quay.io/repository/go-skynet/local-ai?tab=tags
# Image list (dockerhub): https://hub.docker.com/r/localai/localai
image: quay.io/go-skynet/local-ai:master
build:
context: .
dockerfile: Dockerfile
args:
- IMAGE_TYPE=core
- BASE_IMAGE=ubuntu:24.04
ports:
- 8080:8080
env_file:
- .env
environment:
- MODELS_PATH=/models
# Avoid probing remote gallery GGUF metadata during container startup.
# Remove this line or set a positive limit to opt back into cache warming.
- LOCALAI_VRAM_WARM_LIMIT=0
# - DEBUG=true
## Agents (LocalAGI) - https://localai.io/features/agents/
# - LOCALAI_DISABLE_AGENTS=false
# - LOCALAI_AGENT_POOL_DEFAULT_MODEL=hermes-3-llama3.1-8b
# - LOCALAI_AGENT_POOL_ENABLE_SKILLS=true
# - LOCALAI_AGENT_POOL_ENABLE_LOGS=true
# - LOCALAI_AGENT_HUB_URL=https://agenthub.localai.io
## Uncomment to use PostgreSQL for the knowledge base (requires the postgres service below)
# - LOCALAI_AGENT_POOL_VECTOR_ENGINE=postgres
# - LOCALAI_AGENT_POOL_DATABASE_URL=postgresql://localrecall:localrecall@postgres:5432/localrecall?sslmode=disable
volumes:
- models:/models
- images:/tmp/generated/images/
- data:/data
- backends:/backends
- configuration:/configuration
command:
# Here we can specify a list of models to run (see quickstart https://localai.io/basics/getting_started/#running-models )
# or an URL pointing to a YAML configuration file, for example:
# - https://gist.githubusercontent.com/mudler/ad601a0488b497b69ec549150d9edd18/raw/a8a8869ef1bb7e3830bf5c0bae29a0cce991ff8d/phi-2.yaml
- phi-2-chat
# For NVIDIA GPU support with CDI (recommended for NVIDIA Container Toolkit 1.14+):
# Uncomment the following deploy section and use driver: nvidia.com/gpu.
# Include `utility` in capabilities so nvidia-smi / NVML are available —
# without it, free-VRAM reporting on discrete GPUs is unavailable and the
# Nodes UI will misreport memory usage.
# environment:
# NVIDIA_DRIVER_CAPABILITIES: "compute,utility"
# init: true # avoids zombie-reap races that can make nvidia-smi flaky
# deploy:
# resources:
# reservations:
# devices:
# - driver: nvidia.com/gpu
# count: all
# capabilities: [gpu, utility]
#
# For legacy NVIDIA driver (for older NVIDIA Container Toolkit):
# Request compute for CUDA libraries (libcuda.so.1) and utility for NVML.
# environment:
# NVIDIA_DRIVER_CAPABILITIES: "compute,utility"
# init: true
# deploy:
# resources:
# reservations:
# devices:
# - driver: nvidia
# count: 1
# capabilities: [gpu, compute, utility]
## Uncomment for PostgreSQL-backed knowledge base (see Agents docs)
# postgres:
# image: quay.io/mudler/localrecall:v0.5.2-postgresql
# environment:
# - POSTGRES_DB=localrecall
# - POSTGRES_USER=localrecall
# - POSTGRES_PASSWORD=localrecall
# volumes:
# - postgres_data:/var/lib/postgresql
# healthcheck:
# test: ["CMD-SHELL", "pg_isready -U localrecall"]
# interval: 10s
# timeout: 5s
# retries: 5
volumes:
models:
images:
data:
configuration:
backends:
# postgres_data:
This file may be out of date. Check the LocalAI documentation for current setup steps.
Tags
Track your self-hosted stack
Bookmark software to try, rate tools you've used, and keep your collection in one place.
Metadata extracted from README on Oct 6, 2026
Related Software
SillyTavern
34.3k
SillyTavern is a locally installed interface for interacting with text generation LLMs, image generation engines, and TTS voice models.
Artificial IntelligenceAI InterfacesCopyleft License+3
Single Container
text-generation-webui
47.7k
A desktop app for running local LLMs with chat, vision, tool-calling, training, image generation, and an API.
Artificial IntelligenceSelf Hosting SolutionsAI Interfaces+2
Single Container

LibreChat
45.5k
LibreChat is a self-hosted AI chat platform that unifies major AI providers in a ChatGPT-like interface.
Artificial IntelligenceAI InterfacesPermissive License+4
5 Containers
big-AGI
7.1k
Big-AGI is a multi-model AI workspace for experts, focused on chat, model comparison, and local-first self-hosting.
Artificial IntelligenceAI InterfacesAI Media Generation+2
Open Core
AnythingLLM
66.9k
An all-in-one AI application for chatting with documents, building AI agents, and running local or cloud LLM workflows.
Artificial IntelligenceAutomation ToolsSelf Hosting Solutions+7
Single Container
Perplexica
37.2k
Vane is a privacy-focused AI answering engine that runs on your own hardware and returns cited answers from web search and local or cloud LLMs.
Artificial IntelligencePermissive LicenseSearch Engines+5
Vane
37.2k
Vane is a privacy-focused AI answering engine that runs on your own hardware and combines web search with local or cloud LLMs to deliver cited answers.
Artificial IntelligencePermissive LicenseGenerative AI+6
Local Deep Research
9.2k
An AI-powered research assistant for deep, agentic research with citations, designed to run locally or via a self-hosted setup.
Artificial IntelligenceAI RAG SystemsAI Agents+6
3 Containers
IronClaw
12.6k
A secure personal AI assistant that runs with local data storage, sandboxed tools, and multiple interaction channels.
Artificial IntelligenceClaws & ClankersAI Agents+3
Khoj
37.6k
Khoj is a personal AI app for chat, search, and automation that can run locally or as a cloud service.
Artificial IntelligenceCopyleft LicenseGenerative AI+5
5 ContainersOpen Core
machtiani
191
Machtiani (mct) is a terminal-based local code chat service for working with large real codebases and syncing project history for context-aware answers.
Artificial IntelligenceAI Coding AssistantAI Agents+3
BionicGPT
2.4k
Bionic is an on-premise, ChatGPT-like generative AI platform for running locally or at data-center scale with strict data confidentiality.
Artificial IntelligenceAI InterfacesAI RAG Systems+3
Open Core
