Updated FP8 version of Qwen3-30B-A3B non-thinking mode, with better tool use, coding, instruction following, logical reasoning and text comprehension capabilities
On-demand DeploymentDocs | On-demand deployments allow you to use Qwen3 30B A3B Instruct 2507 on dedicated GPUs with Fireworks' high-performance serving stack with high reliability and no rate limits. |