Qwen2.5 are a series of decoder-only language models developed by Qwen team, Alibaba Cloud, available in 0.5B, 1.5B, 3B, 7B, 14B, 32B, and 72B sizes, and base and instruct variants.
Qwen2.5 32B Instruct can be customized with your data to improve responses. Fireworks uses LoRA to efficiently train and deploy your personalized model
On-demand deployments allow you to use Qwen2.5 32B Instruct on dedicated GPUs with Fireworks' high-performance serving stack with high reliability and no rate limits.