Join us for our inaugural conference, Forge 2026

Model Library
/Qwen/Qwen 3.8 Max
model path:accounts/fireworks/models/qwen3p8-max

Built on the architectural foundation of Qwen3.5, Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. Beyond answering harder questions, Qwen3.8 is designed to carry complex, multi-step tasks through to completion with greater reliability.

Qwen 3.8 Max API Features

Serverless

Docs

Qwen 3.8 Max is available via Fireworks' serverless API, where you pay per token. There are several ways to call the Fireworks API, including Fireworks' Python client, the REST API, or OpenAI's Python client.

On-demand Deployment

Docs

On-demand deployments allow you to use Qwen 3.8 Max on dedicated GPUs with Fireworks' high-performance serving stack with high reliability and no rate limits.

Available Serverless

Run queries immediately, pay only for usage

$2.00 / $0.25 / $6.00
Per 1M Tokens (input/cached input/output)

Metadata

State
Ready
Created on
8/5/2026
Kind
Base model
Provider
Qwen

Specification

Calibrated
No
Mixture-of-Experts
Yes

Supported Functionality

Fine-tuning
Not supported
Serverless
Supported
Context Length
N/A
Function Calling
Supported
Embeddings
Not supported
Rerankers
Not supported
Support image input
Supported