Voltar ao guia
RUNTIME / EASY

Ollama

One-command local chat and app integration

Fonte primária
DifficultyEasySetup profile
Platforms5cpu · apple · nvidia · amd · intel
APIOpenAI-compatible + native API:11434
01 / INSTALL & RUN
  1. 1
    Installcurl -fsSL https://ollama.com/install.sh | sh
  2. 2
    Start a modelollama run google/gemma-3-1b-it
  3. 3
    Connect your app

    OpenAI-compatible + native API · :11434. Keep the service bound to localhost unless you add authentication and network controls.

Atenção

GPU and driver support varies by operating system. Confirm the current hardware matrix before buying hardware.

02 / COMPATIBLE MODELS

Modelos open-weight representativos

Google

Gemma 3 1B

1BINT4 / GGUF Q4
Memória estimada
~1.4 GB
Contexto
32K
Meta

Llama 3.2 3B

3BGGUF Q4
Memória estimada
~2.8 GB
Contexto
128K
Alibaba Qwen

Qwen3 4B

4BGGUF Q4
Memória estimada
~3.6 GB
Contexto
32K+
Google

Gemma 3 4B

4BINT4 / GGUF Q4
Memória estimada
~4.2 GB
Contexto
128K
Alibaba Qwen

Qwen3 8B

8BGGUF Q4
Memória estimada
~6.8 GB
Contexto
32K+
Google

Gemma 3 12B

12BINT4
Memória estimada
~9.4 GB
Contexto
128K