selfhostedworld.com logoselfhostedworld.com

Try describing what you need:

LocalAI logo

LocalAI

★ 49.5k

A self-hosted alternative to Chatgpt

MIT LicenseOpen Source — No PaywallSingle Container — single container — minimal hosting footprintDocker · Binary · SourceNo external DB17+ active committers (12mo)

About

LocalAI is an open-source AI engine designed for self-hosted use. It lets users run a broad range of model types locally, including text, vision, speech, image, and video workloads, while keeping...

Community Ratings

No ratings yet

Features

  • OpenAI-compatible APIs
  • Anthropic API compatibility
  • ElevenLabs API compatibility
  • LLMs, vision, voice, image, and video support
  • On-demand backend loading
  • Multi-user API key auth
  • User quotas
  • Role-based access
  • Built-in AI agents
  • Terminal agent

Details

Last Updated
Oct 10, 2026
Created
Mar 18, 2023
Install Methods
dockerbinarysource
Requirements
macOS (for the DMG quickstart)Docker or Podman for container-based setupGPU hardware is optional; CPU-only is supported
Authentication
local
API
REST
Runtime / Stack
DockerK8S
Privacy & Independence
No Cloud RequiredOffline Capable
User Model
Multi-User
Deployment

Deployment: Docker Compose ✓Single Container — minimal hosting footprintsingle container

Run LocalAI with Docker Compose

Compose file from the LocalAI project repository.

1 service

no resource limits declared

needs persistent storage (5 volumes)

docker-compose.yaml · @ 61f4f67 · scanned Sep 28, 2026

services:
  api:
    # See https://localai.io/basics/getting_started/#container-images for
    # a list of available container images (or build your own with the provided Dockerfile)
    # Available images with CUDA, ROCm, SYCL
    # Image list (quay.io): https://quay.io/repository/go-skynet/local-ai?tab=tags
    # Image list (dockerhub): https://hub.docker.com/r/localai/localai
    image: quay.io/go-skynet/local-ai:master
    build:
      context: .
      dockerfile: Dockerfile
      args:
      - IMAGE_TYPE=core
      - BASE_IMAGE=ubuntu:24.04
    ports:
      - 8080:8080
    env_file:
      - .env
    environment:
      - MODELS_PATH=/models
      # Avoid probing remote gallery GGUF metadata during container startup.
      # Remove this line or set a positive limit to opt back into cache warming.
      - LOCALAI_VRAM_WARM_LIMIT=0
    #  - DEBUG=true
    ## Agents (LocalAGI) - https://localai.io/features/agents/
    #  - LOCALAI_DISABLE_AGENTS=false
    #  - LOCALAI_AGENT_POOL_DEFAULT_MODEL=hermes-3-llama3.1-8b
    #  - LOCALAI_AGENT_POOL_ENABLE_SKILLS=true
    #  - LOCALAI_AGENT_POOL_ENABLE_LOGS=true
    #  - LOCALAI_AGENT_HUB_URL=https://agenthub.localai.io
    ## Uncomment to use PostgreSQL for the knowledge base (requires the postgres service below)
    #  - LOCALAI_AGENT_POOL_VECTOR_ENGINE=postgres
    #  - LOCALAI_AGENT_POOL_DATABASE_URL=postgresql://localrecall:localrecall@postgres:5432/localrecall?sslmode=disable
    volumes:
      - models:/models
      - images:/tmp/generated/images/
      - data:/data
      - backends:/backends
      - configuration:/configuration
    command:
    # Here we can specify a list of models to run (see quickstart https://localai.io/basics/getting_started/#running-models )
    # or an URL pointing to a YAML configuration file, for example:
    # - https://gist.githubusercontent.com/mudler/ad601a0488b497b69ec549150d9edd18/raw/a8a8869ef1bb7e3830bf5c0bae29a0cce991ff8d/phi-2.yaml
    - phi-2-chat
    # For NVIDIA GPU support with CDI (recommended for NVIDIA Container Toolkit 1.14+):
    # Uncomment the following deploy section and use driver: nvidia.com/gpu.
    # Include `utility` in capabilities so nvidia-smi / NVML are available —
    # without it, free-VRAM reporting on discrete GPUs is unavailable and the
    # Nodes UI will misreport memory usage.
    # environment:
    #   NVIDIA_DRIVER_CAPABILITIES: "compute,utility"
    # init: true   # avoids zombie-reap races that can make nvidia-smi flaky
    # deploy:
    #   resources:
    #     reservations:
    #       devices:
    #         - driver: nvidia.com/gpu
    #           count: all
    #           capabilities: [gpu, utility]
    #
    # For legacy NVIDIA driver (for older NVIDIA Container Toolkit):
    # Request compute for CUDA libraries (libcuda.so.1) and utility for NVML.
    # environment:
    #   NVIDIA_DRIVER_CAPABILITIES: "compute,utility"
    # init: true
    # deploy:
    #   resources:
    #     reservations:
    #       devices:
    #         - driver: nvidia
    #           count: 1
    #           capabilities: [gpu, compute, utility]

  ## Uncomment for PostgreSQL-backed knowledge base (see Agents docs)
  # postgres:
  #   image: quay.io/mudler/localrecall:v0.5.2-postgresql
  #   environment:
  #     - POSTGRES_DB=localrecall
  #     - POSTGRES_USER=localrecall
  #     - POSTGRES_PASSWORD=localrecall
  #   volumes:
  #     - postgres_data:/var/lib/postgresql
  #   healthcheck:
  #     test: ["CMD-SHELL", "pg_isready -U localrecall"]
  #     interval: 10s
  #     timeout: 5s
  #     retries: 5

volumes:
  models:
  images:
  data:
  configuration:
  backends:
  # postgres_data:

This file may be out of date. Check the LocalAI documentation for current setup steps.

Track your self-hosted stack

Bookmark software to try, rate tools you've used, and keep your collection in one place.

Metadata extracted from README on Oct 6, 2026

Related Software

SillyTavern logo

SillyTavern

34.3k
SillyTavern is a locally installed interface for interacting with text generation LLMs, image generation engines, and TTS voice models.
Artificial IntelligenceAI InterfacesCopyleft License+3
Single Container
Details
text-generation-webui logo

text-generation-webui

47.7k
A desktop app for running local LLMs with chat, vision, tool-calling, training, image generation, and an API.
Artificial IntelligenceSelf Hosting SolutionsAI Interfaces+2
Single Container
Details
LibreChat logo

LibreChat

45.5k
LibreChat is a self-hosted AI chat platform that unifies major AI providers in a ChatGPT-like interface.
Artificial IntelligenceAI InterfacesPermissive License+4
5 Containers
Details
big-AGI logo

big-AGI

7.1k
Big-AGI is a multi-model AI workspace for experts, focused on chat, model comparison, and local-first self-hosting.
Artificial IntelligenceAI InterfacesAI Media Generation+2
Open Core
Details
AnythingLLM logo

AnythingLLM

66.9k
An all-in-one AI application for chatting with documents, building AI agents, and running local or cloud LLM workflows.
Artificial IntelligenceAutomation ToolsSelf Hosting Solutions+7
Single Container
Details
Perplexica logo

Perplexica

37.2k
Vane is a privacy-focused AI answering engine that runs on your own hardware and returns cited answers from web search and local or cloud LLMs.
Artificial IntelligencePermissive LicenseSearch Engines+5
Details
Vane logo

Vane

37.2k
Vane is a privacy-focused AI answering engine that runs on your own hardware and combines web search with local or cloud LLMs to deliver cited answers.
Artificial IntelligencePermissive LicenseGenerative AI+6
Details
Local Deep Research logo

Local Deep Research

9.2k
An AI-powered research assistant for deep, agentic research with citations, designed to run locally or via a self-hosted setup.
Artificial IntelligenceAI RAG SystemsAI Agents+6
3 Containers
Details
IronClaw logo

IronClaw

12.6k
A secure personal AI assistant that runs with local data storage, sandboxed tools, and multiple interaction channels.
Artificial IntelligenceClaws & ClankersAI Agents+3
Details
Khoj logo

Khoj

37.6k
Khoj is a personal AI app for chat, search, and automation that can run locally or as a cloud service.
Artificial IntelligenceCopyleft LicenseGenerative AI+5
5 ContainersOpen Core
Details
machtiani logo

machtiani

191
Machtiani (mct) is a terminal-based local code chat service for working with large real codebases and syncing project history for context-aware answers.
Artificial IntelligenceAI Coding AssistantAI Agents+3
Details
BionicGPT logo

BionicGPT

2.4k
Bionic is an on-premise, ChatGPT-like generative AI platform for running locally or at data-center scale with strict data confidentiality.
Artificial IntelligenceAI InterfacesAI RAG Systems+3
Open Core
Details