DeepSeek-V4-Pro-0813 available now on Fireworks

Fireworks Blog

Header Image Do Open Models Perform Silent Reasoning Before They Respond

Can open models carry readable silent signals before they speak? Reproducing J-Lens Readouts on Kimi K3 & Qwen3.5-9B

Llama 4 Maverick on Fireworks
Developer Experience
4/28/2025

Optimizing Llama 4 Maverick on Fireworks

RAG application using MongoDB Atlas and Fireworks
Developer Experience
4/9/2025

Building Enterprise-Scale RAG Systems with Fireworks and MongoDB Atlas

Fireworks Now Supports NVIDIA NIM Deployments for Blazing AI Inference
Model Releases
3/18/2025

Fireworks Now Supports NVIDIA NIM Deployments for Blazing AI Inference

Faster, more efficient DeepSeek on the Fireworks Developer Cloud
Model Releases
3/18/2025

Faster, more efficient DeepSeek on the Fireworks Developer Cloud

Fine-Tuning DeepSeek v3 & R1 to optimize quality, latency, & cost
Model Releases
3/12/2025

Fine-Tuning DeepSeek v3 & R1 to optimize quality, latency, & cost

Enabling Function Calling in DeepSeek v3: Bridging the Gap Between Text and Action
Model Releases
2/14/2025

Enabling Function Calling in DeepSeek v3: Bridging the Gap Between Text and Action

DeepSeek v3 and R1 Model Architecture: Why it's powerful and economical
Developer Experience
2/7/2025

DeepSeek v3 and R1 Model Architecture: Why it's powerful and economical

DeepSeek R1 Just Got Eyes with Fireworks Document Inlining
Model Releases
2/5/2025

DeepSeek R1 Just Got Eyes with Fireworks Document Inlining

From text to task: Constrained generation for structured extraction in R1
Developer Experience
2/1/2025

From text to task: Constrained generation for structured extraction in R1

Distillation with Reasoning: Can DeepSeek R1 Teach Better Than Humans?
Developer Experience
1/31/2025

Distillation with Reasoning: Can DeepSeek R1 Teach Better Than Humans?

Mistral Small 3 Now Available on Fireworks: Faster, Lighter, and More Efficient
Model Releases
1/30/2025

Mistral Small 3 Now Available on Fireworks: Faster, Lighter, and More Efficient

Beyond Supervised Fine Tuning: How Reinforcement Learning Empowers AI with Minimal Labels
Developer Experience
1/27/2025

Beyond Supervised Fine Tuning: How Reinforcement Learning Empowers AI with Minimal Labels

DeepSeek R1: All you need to know 🐳
Model Releases
1/24/2025

DeepSeek R1: All you need to know 🐳

Real-time, performant code assistance: How Sourcegraph scaled with Fireworks
Case Studies
1/22/2025

Real-time, performant code assistance: How Sourcegraph scaled with Fireworks

DeepSeek V3 just got vision capabilities!
Model Releases
12/18/2024

DeepSeek V3 just got vision capabilities!

20x faster Whisper than OpenAI - Fireworks audio transcribes 1 hour in 4 seconds
Model Releases
12/9/2024

20x faster Whisper than OpenAI - Fireworks audio transcribes 1 hour in 4 seconds

How Cresta drives millions of real-time, AI-powered contact center interactions with Fireworks
Case Studies
12/8/2024

How Cresta drives millions of real-time, AI-powered contact center interactions with Fireworks

Fireworks f1: A breakthrough in complex reasoning with Compound AI
Model Releases
11/15/2024

Fireworks f1: A breakthrough in complex reasoning with Compound AI

How Upwork and Fireworks deliver faster, smarter proposals for freelancers
Case Studies
11/11/2024

How Upwork and Fireworks deliver faster, smarter proposals for freelancers

FLUX.1 on Fireworks: Fast, frugal, and flexible
Model Releases
10/22/2024

FLUX.1 on Fireworks: Fast, frugal, and flexible

FireAttention V3: Enabling AMD as a viable alternative for GPU inference
Developer Experience
10/15/2024

FireAttention V3: Enabling AMD as a viable alternative for GPU inference