Join us for our inaugural conference, Forge 2026

Fireworks Blog

Phylo and Fireworks

Phylo brings frontier AI to more scientists with open models on Fireworks

Phylo cut inference cost 60% while doubling users month-on-month, running Biomni Lab's long-horizon biology agents on open models on Fireworks.

Updated Supervised Fine Tuning
Model Releases
6/13/2025

Introducing Supervised Fine Tuning V2

Reinforcement fine tuning announcement
Model Releases
6/9/2025

Reinforcement Fine Tuning (Beta): Train expert open models to surpass closed frontier models

Fireworks Dev Day 2025 Wrapped
Company News
5/29/2025

Fireworks DevDay 2025 Wrapped

 Independent benchmarking of Fireworks shows >250 tokens / second on DeepSeek V3
Model Releases
5/28/2025

FireAttention V4: Industry-Leading Latency and Cost Efficiency with FP4

Building an open-source Browser Agent on Fireworks
Developer Experience
5/21/2025

Building an open-source Browser Agent on Fireworks

Agentic AI Systems
Developer Experience
5/19/2025

Agentic AI Systems

Supervised Fine-Tuning (SFT) with LoRA on Fireworks: Tutorial
Developer Experience
5/12/2025

Supervised Fine-Tuning (SFT) with LoRA on Fireworks: Tutorial

Qwen 3 on Fireworks
Model Releases
5/6/2025

Qwen 3 on Fireworks: Controllable Chain-of-Thought and Tool Calling at Frontier Scale

Llama 4 Maverick on Fireworks
Developer Experience
4/28/2025

Optimizing Llama 4 Maverick on Fireworks

RAG application using MongoDB Atlas and Fireworks
Developer Experience
4/9/2025

Building Enterprise-Scale RAG Systems with Fireworks and MongoDB Atlas

Fireworks Now Supports NVIDIA NIM Deployments for Blazing AI Inference
Model Releases
3/18/2025

Fireworks Now Supports NVIDIA NIM Deployments for Blazing AI Inference

Faster, more efficient DeepSeek on the Fireworks Developer Cloud
Model Releases
3/18/2025

Faster, more efficient DeepSeek on the Fireworks Developer Cloud

Fine-Tuning DeepSeek v3 & R1 to optimize quality, latency, & cost
Model Releases
3/12/2025

Fine-Tuning DeepSeek v3 & R1 to optimize quality, latency, & cost

Enabling Function Calling in DeepSeek v3: Bridging the Gap Between Text and Action
Model Releases
2/14/2025

Enabling Function Calling in DeepSeek v3: Bridging the Gap Between Text and Action

DeepSeek v3 and R1 Model Architecture: Why it's powerful and economical
Developer Experience
2/7/2025

DeepSeek v3 and R1 Model Architecture: Why it's powerful and economical

DeepSeek R1 Just Got Eyes with Fireworks Document Inlining
Model Releases
2/5/2025

DeepSeek R1 Just Got Eyes with Fireworks Document Inlining

From text to task: Constrained generation for structured extraction in R1
Developer Experience
2/1/2025

From text to task: Constrained generation for structured extraction in R1

Distillation with Reasoning: Can DeepSeek R1 Teach Better Than Humans?
Developer Experience
1/31/2025

Distillation with Reasoning: Can DeepSeek R1 Teach Better Than Humans?

Mistral Small 3 Now Available on Fireworks: Faster, Lighter, and More Efficient
Model Releases
1/30/2025

Mistral Small 3 Now Available on Fireworks: Faster, Lighter, and More Efficient

Beyond Supervised Fine Tuning: How Reinforcement Learning Empowers AI with Minimal Labels
Developer Experience
1/27/2025

Beyond Supervised Fine Tuning: How Reinforcement Learning Empowers AI with Minimal Labels

DeepSeek R1: All you need to know 🐳
Model Releases
1/24/2025

DeepSeek R1: All you need to know 🐳