audio model · orpheus · Windows

OR Can I run Orpheus 3B on Nvidia GeForce RTX 4070 (12GB)?

Compatibility verdict memory check

Yes, it runsfast on this GPU

Yes. Orpheus 3B runs on Nvidia GeForce RTX 4070 (12GB) at Q4_K_M GGUF (~4 GB of ~11 GB usable).

usable ~11 GB

Needs ~4 GB Device usable ~11 GB

Runs at Q4_K_M GGUF using ~4 GB of ~11 GB usable.

How to run it

Use llama.cpp or LM Studio at Q4_K_M GGUF. It is light enough to run on CPU; a GPU just makes it faster.

Model orpheus

Device Windows

You could also run

Run Orpheus 3B on other hardware

FAQ

Can Nvidia GeForce RTX 4070 (12GB) run Orpheus 3B?

Yes. Orpheus 3B runs on Nvidia GeForce RTX 4070 (12GB) at Q4_K_M GGUF (~4 GB of ~11 GB usable).

How much memory does Orpheus 3B need?

Nvidia GeForce RTX 4070 (12GB) has room to spare. At Q4_K_M GGUF the realistic peak is ~4 GB of memory.

What do I use to run Orpheus 3B locally?

Orpheus 3B runs in llama.cpp or LM Studio (among others). It runs on CPU, so no GPU is required.

Sources

VRAM figures are sourced peak-usage anchors at the noted quant, validated 2026-06-15. See methodology.