Kembali
RUNTIME / MEDIUM

LocalAI

One OpenAI-compatible endpoint for text, vision, speech and image models, GPU optional

Sumber primer
DifficultyMediumSetup profile
Platforms5cpu · apple · nvidia · amd · intel
APIOpenAI + Anthropic + ElevenLabs-compatible:8080
01 / INSTALL & RUN

Installation depends on your OS. These commands are examples or templates: replace placeholders and verify quantization, file format, drivers and runtime support. Memory fit does not guarantee deployment.

  1. 1
    Installdocker run -ti --name local-ai -p 8080:8080 localai/localai:latest
  2. 2
    Start a modellocal-ai run <verified-gallery-or-Hugging-Face-model-name>
  3. 3
    Connect your app

    OpenAI + Anthropic + ElevenLabs-compatible · :8080. Keep the service bound to localhost unless you add authentication and network controls.

Perhatian

Backends are separate images pulled on demand, so the first run needs network access and disk. Pick the image tag for your accelerator; the CPU image does not silently enable a GPU.

02 / COMPATIBLE MODELS

Model open-weight perwakilan

Alibaba Qwen

Qwen3.6 27B

27BQ4 / NVFP4 estimate
Estimasi memori
~18.5 GB
Konteks
262K native
Alibaba Qwen

Qwen3.8 27B

27BQ4 estimate
Estimasi memori
~19.5 GB
Konteks
262K native
Alibaba Qwen

Qwen3.6 35B-A3B

35B3B active · Q4 / NVFP4 MoE
Estimasi memori
~25 GB
Konteks
262K native
DeepSeek

DeepSeek V4 Flash

284B13B active · Official FP4 + FP8 mixed
Estimasi memori
~176 GB
Konteks
1M native
Google

Gemma 3 1B

1BINT4 / GGUF Q4
Estimasi memori
~1.4 GB
Konteks
32K
Meta

Llama 3.2 3B

3BGGUF Q4
Estimasi memori
~2.8 GB
Konteks
128K