DeepSeek-V4-Pro-0813 available now on Fireworks

Model Library
/Fireworks/Nemotron 3 Embed 8B
Fireworks Logo Mark

Nemotron 3 Embed 8B

Ready
model path:accounts/fireworks/models/nemotron-3-embed-8b-bf16

NVIDIA Nemotron-3-Embed-8B - an 8B-parameter text embedding model producing 4096-dimensional embeddings, with matryoshka support for truncating to smaller dimensions. Well suited for retrieval, semantic search, and RAG workloads.

Nemotron 3 Embed 8B API Features

On-demand Deployment

Docs

On-demand deployments allow you to use Nemotron 3 Embed 8B on dedicated GPUs with Fireworks' high-performance serving stack with high reliability and no rate limits.

Metadata

State
Ready
Created on
7/24/2026
Kind
Embedding model
Provider
Fireworks

Specification

Calibrated
No
Mixture-of-Experts
No
Parameters
8B

Supported Functionality

Fine-tuning
Not supported
Serverless
Not supported
Context Length
32.7k tokens
Function Calling
Not supported
Embeddings
Supported
Rerankers
Not supported
Support image input
Not supported