HARDWARE / SYSTEM RAM
8 GB CPU laptop
Small text models with short context. Expect patient, private inference rather than speed.
Everyday laptop8 GB
01 / PRACTICAL CEILING
Qwen3 4B
The highest listed result may be a tight or offloaded fit. For daily use, prefer the first result marked Comfortable and keep context modest.
02 / DEFAULT RUNTIME
llama.cpp
GGUF models, CPU/GPU offload, embedded and unusual hardware. New architectures may require a recent build. Backend availability does not imply equal performance.
가이드 열기03 / MODEL ENVELOPE
대표 오픈 웨이트 모델
Alibaba Qwen
Qwen3 4B
4BGGUF Q4
- 추정 메모리
- ~3.6 GB
- 컨텍스트
- 32K+
Google
Gemma 3 4B
4BINT4 / GGUF Q4
- 추정 메모리
- ~4.2 GB
- 컨텍스트
- 128K
Meta
Llama 3.2 3B
3BGGUF Q4
- 추정 메모리
- ~2.8 GB
- 컨텍스트
- 128K
Google
Gemma 3 1B
1BINT4 / GGUF Q4
- 추정 메모리
- ~1.4 GB
- 컨텍스트
- 32K
주의
Fit is based on estimated total model memory. Driver support, KV cache, multimodal projectors, concurrency, and desktop applications can all reduce available headroom. Current fit: Qwen3 4B (Comfortable), Gemma 3 4B (Tight), Llama 3.2 3B (Comfortable), Gemma 3 1B (Comfortable).