가이드로
HARDWARE / UNIFIED MEMORY

Mac Studio M5 Ultra · 96 GB

The entry M5 Ultra configuration combines a 96 GB unified pool with 1.2 TB/s memory bandwidth. Apple announced it in August 2026; availability begins September 22.

1차 출처
Current Apple desktop96 GB
01 / PRACTICAL CEILING

Qwen3.8 Flash Next REAP-288

The highest listed result may be a tight or offloaded fit. For daily use, prefer the first result marked Comfortable and keep context modest.

02 / DEFAULT RUNTIME

MLX LM

Native Apple silicon inference, experimentation and fine-tuning. Only for Apple silicon. Use an MLX-converted model and tune KV cache size for long contexts.

가이드 열기
03 / MODEL ENVELOPE

대표 오픈 웨이트 모델

Community REAP build

Qwen3.8 Flash Next REAP-288

176BREAP-288 · MLX 4-bit · PLE on NVMe
추정 메모리
~40.57 GB
컨텍스트
7.6K recommended on measured 48 GB Mac
Meta

Llama 3.1 70B

70BGGUF Q4
추정 메모리
~46 GB
컨텍스트
128K
Alibaba Qwen

Qwen3.6 35B-A3B

35B3B active · Q4 / NVFP4 MoE
추정 메모리
~25 GB
컨텍스트
262K native
Alibaba Qwen

Qwen3 32B

32BGGUF Q4
추정 메모리
~24.5 GB
컨텍스트
32K+
Alibaba Qwen

Qwen3 30B-A3B

30B3B active · GGUF Q4 MoE
추정 메모리
~22.5 GB
컨텍스트
32K+
Alibaba Qwen

Qwen3.6 27B

27BQ4 / NVFP4 estimate
추정 메모리
~18.5 GB
컨텍스트
262K native
주의

Fit is based on estimated total model memory. Driver support, KV cache, multimodal projectors, concurrency, and desktop applications can all reduce available headroom. Current fit: Qwen3.8 Flash Next REAP-288 (Comfortable), Llama 3.1 70B (Comfortable), Qwen3.6 35B-A3B (Comfortable), Qwen3 32B (Comfortable), Qwen3 30B-A3B (Comfortable), Qwen3.6 27B (Comfortable).