RUNTIME / MEDIUM
LocalAI
One OpenAI-compatible endpoint for text, vision, speech and image models, GPU optional
一次情報DifficultyMediumSetup profile
Platforms5cpu · apple · nvidia · amd · intel
APIOpenAI + Anthropic + ElevenLabs-compatible:8080
01 / INSTALL & RUN
Installation depends on your OS. These commands are examples or templates: replace placeholders and verify quantization, file format, drivers and runtime support. Memory fit does not guarantee deployment.
- 1Install
docker run -ti --name local-ai -p 8080:8080 localai/localai:latest - 2Start a model
local-ai run <verified-gallery-or-Hugging-Face-model-name> - 3Connect your app
OpenAI + Anthropic + ElevenLabs-compatible · :8080. Keep the service bound to localhost unless you add authentication and network controls.
注意
Backends are separate images pulled on demand, so the first run needs network access and disk. Pick the image tag for your accelerator; the CPU image does not silently enable a GPU.
02 / COMPATIBLE MODELS
代表的なオープンウェイトモデル
Alibaba Qwen
Qwen3.6 27B
27BQ4 / NVFP4 estimate
- 推定メモリ
- ~18.5 GB
- コンテキスト
- 262K native
Alibaba Qwen
Qwen3.8 27B
27BQ4 estimate
- 推定メモリ
- ~19.5 GB
- コンテキスト
- 262K native
Alibaba Qwen
Qwen3.6 35B-A3B
35B3B active · Q4 / NVFP4 MoE
- 推定メモリ
- ~25 GB
- コンテキスト
- 262K native
DeepSeek
DeepSeek V4 Flash
284B13B active · Official FP4 + FP8 mixed
- 推定メモリ
- ~176 GB
- コンテキスト
- 1M native
Google
Gemma 3 1B
1BINT4 / GGUF Q4
- 推定メモリ
- ~1.4 GB
- コンテキスト
- 32K
Meta
Llama 3.2 3B
3BGGUF Q4
- 推定メモリ
- ~2.8 GB
- コンテキスト
- 128K