Back to guide
RUNTIME / EASY

Ollama

One-command local chat and app integration

Primary source
DifficultyEasySetup profile
Platforms5cpu · apple · nvidia · amd · intel
APIOpenAI-compatible + native API:11434
01 / INSTALL & RUN
  1. 1
    Installcurl -fsSL https://ollama.com/install.sh | sh
  2. 2
    Start a modelollama run google/gemma-3-1b-it
  3. 3
    Connect your app

    OpenAI-compatible + native API · :11434. Keep the service bound to localhost unless you add authentication and network controls.

Watch out

GPU and driver support varies by operating system. Confirm the current hardware matrix before buying hardware.

02 / COMPATIBLE MODELS

Representative open-weight models

Google

Gemma 3 1B

1BINT4 / GGUF Q4
Estimated memory
~1.4 GB
Context
32K
Meta

Llama 3.2 3B

3BGGUF Q4
Estimated memory
~2.8 GB
Context
128K
Alibaba Qwen

Qwen3 4B

4BGGUF Q4
Estimated memory
~3.6 GB
Context
32K+
Google

Gemma 3 4B

4BINT4 / GGUF Q4
Estimated memory
~4.2 GB
Context
128K
Alibaba Qwen

Qwen3 8B

8BGGUF Q4
Estimated memory
~6.8 GB
Context
32K+
Google

Gemma 3 12B

12BINT4
Estimated memory
~9.4 GB
Context
128K