Retour au guide
RUNTIME / EASY

Ollama

One-command local chat and app integration

Source primaire
DifficultyEasySetup profile
Platforms5cpu · apple · nvidia · amd · intel
APIOpenAI-compatible + native API:11434
01 / INSTALL & RUN
  1. 1
    Installcurl -fsSL https://ollama.com/install.sh | sh
  2. 2
    Start a modelollama run google/gemma-3-1b-it
  3. 3
    Connect your app

    OpenAI-compatible + native API · :11434. Keep the service bound to localhost unless you add authentication and network controls.

Attention

GPU and driver support varies by operating system. Confirm the current hardware matrix before buying hardware.

02 / COMPATIBLE MODELS

Modèles open-weight représentatifs

Google

Gemma 3 1B

1BINT4 / GGUF Q4
Mémoire estimée
~1.4 GB
Contexte
32K
Meta

Llama 3.2 3B

3BGGUF Q4
Mémoire estimée
~2.8 GB
Contexte
128K
Alibaba Qwen

Qwen3 4B

4BGGUF Q4
Mémoire estimée
~3.6 GB
Contexte
32K+
Google

Gemma 3 4B

4BINT4 / GGUF Q4
Mémoire estimée
~4.2 GB
Contexte
128K
Alibaba Qwen

Qwen3 8B

8BGGUF Q4
Mémoire estimée
~6.8 GB
Contexte
32K+
Google

Gemma 3 12B

12BINT4
Mémoire estimée
~9.4 GB
Contexte
128K