MODEL / COMMUNITY REAP BUILD
Qwen3.8 Flash Next REAP-288
Community-pruned conversion, not the stock checkpoint. The 39 GB model / ~40.57 GB total resident path requires oMLX 0.6.4+ and streams only the sparse PLE table from NVMe. A published M4 Pro 48 GB test measured 26–29 tok/s and recommended 7,680 tokens; do not generalize it to arbitrary models or Macs.
一手來源總參數176B
估算記憶體~40.57 GBREAP-288 · MLX 4-bit · PLE on NVMe
上下文7.6K recommended on measured 48 GB MacText
LicenseQwen CommunityMultilingual; post-pruning quality not independently established
01 / 適合
- 48 GB Mac experiments
- coding agents
- model-specific NVMe streaming
02 / 估算適配
03 / 部署方式
oMLXAdvanced
omlx serve --model-dir ~/models --memory-guard-gb 48注意
Community-pruned conversion, not the stock checkpoint. The 39 GB model / ~40.57 GB total resident path requires oMLX 0.6.4+ and streams only the sparse PLE table from NVMe. A published M4 Pro 48 GB test measured 26–29 tok/s and recommended 7,680 tokens; do not generalize it to arbitrary models or Macs.