RUNTIME / EASY
Lemonade
Ryzen AI, Radeon and Strix Halo PCs, including the XDNA2 NPU, behind one local API
一手来源DifficultyEasySetup profile
Platforms4amd · nvidia · cpu · apple
APIOpenAI + Ollama + Anthropic-compatible:13305
01 / INSTALL & RUN
安装方式取决于操作系统。以下命令是示例或模板,请先替换占位符并核对量化版本、文件格式、驱动和运行时支持;内存适配不保证部署成功。
- 1Install
snap install lemonade-server # or the .msi / .deb / .pkg installer - 2Start a model
lemonade run <verified-model-name> - 3Connect your app
OpenAI + Ollama + Anthropic-compatible · :13305. Keep the service bound to localhost unless you add authentication and network controls.
注意
Built with AMD engineers around Ryzen AI, Radeon and Strix Halo; CUDA, Vulkan, Metal and CPU backends cover other PCs. Run `lemonade backends` on the machine itself — the NPU path needs the vendor driver stack and is not available everywhere.
02 / COMPATIBLE MODELS
代表性开放权重模型
Alibaba Qwen
Qwen3.6 27B
27BQ4 / NVFP4 estimate
- 估算内存
- ~18.5 GB
- 上下文
- 262K native
Alibaba Qwen
Qwen3.8 27B
27BQ4 estimate
- 估算内存
- ~19.5 GB
- 上下文
- 262K native
Alibaba Qwen
Qwen3.6 35B-A3B
35B3B active · Q4 / NVFP4 MoE
- 估算内存
- ~25 GB
- 上下文
- 262K native
DeepSeek
DeepSeek V4 Flash
284B13B active · Official FP4 + FP8 mixed
- 估算内存
- ~176 GB
- 上下文
- 1M native
Google
Gemma 3 1B
1BINT4 / GGUF Q4
- 估算内存
- ~1.4 GB
- 上下文
- 32K
Meta
Llama 3.2 3B
3BGGUF Q4
- 估算内存
- ~2.8 GB
- 上下文
- 128K