ガイドへ戻る
HARDWARE / SYSTEM RAM

8 GB CPU laptop

Small text models with short context. Expect patient, private inference rather than speed.

Everyday laptop8 GB
01 / PRACTICAL CEILING

Qwen3 4B

The highest listed result may be a tight or offloaded fit. For daily use, prefer the first result marked Comfortable and keep context modest.

02 / DEFAULT RUNTIME

llama.cpp

GGUF models, CPU/GPU offload, embedded and unusual hardware. New architectures may require a recent build. Backend availability does not imply equal performance.

ガイドを開く
03 / MODEL ENVELOPE

代表的なオープンウェイトモデル

Alibaba Qwen

Qwen3 4B

4BGGUF Q4
推定メモリ
~3.6 GB
コンテキスト
32K+
Google

Gemma 3 4B

4BINT4 / GGUF Q4
推定メモリ
~4.2 GB
コンテキスト
128K
Meta

Llama 3.2 3B

3BGGUF Q4
推定メモリ
~2.8 GB
コンテキスト
128K
Google

Gemma 3 1B

1BINT4 / GGUF Q4
推定メモリ
~1.4 GB
コンテキスト
32K
注意

Fit is based on estimated total model memory. Driver support, KV cache, multimodal projectors, concurrency, and desktop applications can all reduce available headroom. Current fit: Qwen3 4B (Comfortable), Gemma 3 4B (Tight), Llama 3.2 3B (Comfortable), Gemma 3 1B (Comfortable).