Join the Fireworks Startups Program and unlock credits, expert support, and community to scale fast. Join here

Meta Mark

Llama 3.2 90B Vision Instruct

Instruction-tuned image reasoning model with 90B parameters from Meta. Optimized for visual recognition, image reasoning, captioning, and answering general questions about an image. The model can understand visual data, such as charts and graphs and also bridge the gap between vision and language by generating text to describe images details

Try Model

Fireworks Features

On-demand Deployment

On-demand deployments give you dedicated GPUs for Llama 3.2 90B Vision Instruct using Fireworks' reliable, high-performance system with no rate limits.

Learn More

Info & Pricing

Provider

Meta

Model Type

Vision

Context Length

131072

Pricing Per 1M Tokens

$0.9