Skip to content

audio model · orpheus · Windows

Can I run Orpheus 3B on AMD Ryzen AI Halo (128GB)?

Compatibility verdict memory check
Yes, it runs fast on this GPU

Yes. Orpheus 3B runs on AMD Ryzen AI Halo (128GB) at Q4_K_M GGUF (~4 GB of ~96 GB usable).

Needs ~4 GB Device usable ~96 GB

Runs at Q4_K_M GGUF using ~4 GB of ~96 GB usable.

Peak memory
~4 GB
Usable on device
~96 GB
Device memory
128 GB
Quant
Q4_K_M GGUF
Share on X Share on Reddit

How to run it

Use llama.cpp or LM Studio at Q4_K_M GGUF. It is light enough to run on CPU; a GPU just makes it faster.

Model orpheus
Type
Text to speech
Parameters
3B
Peak memory
~4 GB at Q4_K_M GGUF
License
Apache-2.0
Full Orpheus 3B requirements →
Device Windows
Memory
128 GB unified
Usable for weights
~96 GB
Power draw
~120 W
Best runtime
llama.cpp (Vulkan/ROCm) / LM Studio
Best models for AMD Ryzen AI Halo (128GB) →

You could also run

Run Orpheus 3B on other hardware

FAQ

Can AMD Ryzen AI Halo (128GB) run Orpheus 3B?

Yes. Orpheus 3B runs on AMD Ryzen AI Halo (128GB) at Q4_K_M GGUF (~4 GB of ~96 GB usable).

How much memory does Orpheus 3B need?

AMD Ryzen AI Halo (128GB) has room to spare. At Q4_K_M GGUF the realistic peak is ~4 GB of memory.

What do I use to run Orpheus 3B locally?

Orpheus 3B runs in llama.cpp or LM Studio (among others). It runs on CPU, so no GPU is required.

Sources

VRAM figures are sourced peak-usage anchors at the noted quant, validated 2026-08-03. See methodology.