Join us for our inaugural conference, Forge 2026
Medium-sized reasoning model from Qwen.
On-demand deployments allow you to use QWQ 32B on dedicated GPUs with Fireworks' high-performance serving stack with high reliability and no rate limits.