Kimi K3 on Fireworks: Frontier Intelligence You Can Own

Model Library
/Qwen/Qwen3 Embedding 8B
Quen Logo Mark

Qwen3 Embedding 8B

Ready
model path:accounts/fireworks/models/qwen3-embedding-8b

The Qwen3 Embedding 8B model is the latest proprietary model of the Qwen family, specifically designed for text embedding tasks. This model inherits the exceptional multilingual capabilities, long-text understanding, and reasoning skills building upon the dense foundational models of the Qwen3 series. The model represents significant advancements in multiple text embedding tasks including text retrieval, code retrieval, text classification, text clustering.

Qwen3 Embedding 8B API Features

Serverless

Docs

Qwen3 Embedding 8B is available via Fireworks' serverless API, where you pay per token. There are several ways to call the Fireworks API, including Fireworks' Python client, the REST API, or OpenAI's Python client.

On-demand Deployment

Docs

On-demand deployments allow you to use Qwen3 Embedding 8B on dedicated GPUs with Fireworks' high-performance serving stack with high reliability and no rate limits.

Available Serverless

Run queries immediately, pay only for usage

$0.10
Per 1M Tokens

Qwen3 Embedding 8B FAQs

Metadata

State
Ready
Created on
8/20/2025
Kind
Embedding model
Provider
Qwen

Specification

Calibrated
No
Mixture-of-Experts
No
Parameters
8.18B

Supported Functionality

Fine-tuning
Not supported
Serverless
Supported
Context Length
40.9k tokens
Function Calling
Not supported
Embeddings
Supported
Rerankers
Not supported
Support image input
Not supported