RUNTIME / EASY
Ollama
One-command local chat and app integration
प्राथमिक स्रोतDifficultyEasySetup profile
Platforms5cpu · apple · nvidia · amd · intel
APIOpenAI-compatible + native API:11434
01 / INSTALL & RUN
- 1Install
curl -fsSL https://ollama.com/install.sh | sh - 2Start a model
ollama run google/gemma-3-1b-it - 3Connect your app
OpenAI-compatible + native API · :11434. Keep the service bound to localhost unless you add authentication and network controls.
सावधानी
GPU and driver support varies by operating system. Confirm the current hardware matrix before buying hardware.
02 / COMPATIBLE MODELS
प्रतिनिधि ओपन-वेट मॉडल
Google
Gemma 3 1B
1BINT4 / GGUF Q4
- अनुमानित मेमोरी
- ~1.4 GB
- कॉन्टेक्स्ट
- 32K
Meta
Llama 3.2 3B
3BGGUF Q4
- अनुमानित मेमोरी
- ~2.8 GB
- कॉन्टेक्स्ट
- 128K
Alibaba Qwen
Qwen3 4B
4BGGUF Q4
- अनुमानित मेमोरी
- ~3.6 GB
- कॉन्टेक्स्ट
- 32K+
Google
Gemma 3 4B
4BINT4 / GGUF Q4
- अनुमानित मेमोरी
- ~4.2 GB
- कॉन्टेक्स्ट
- 128K
Alibaba Qwen
Qwen3 8B
8BGGUF Q4
- अनुमानित मेमोरी
- ~6.8 GB
- कॉन्टेक्स्ट
- 32K+
Google
Gemma 3 12B
12BINT4
- अनुमानित मेमोरी
- ~9.4 GB
- कॉन्टेक्स्ट
- 128K