DeepSeek R1 0528, an updated version of the state-of-the-art DeepSeek R1 model, is now available. Try it now!

Meta Mark

Llama 3.2 11B Vision Instruct

Instruction-tuned image reasoning model from Meta with 11B parameters. Optimized for visual recognition, image reasoning, captioning, and answering general questions about an image. The model can understand visual data, such as charts and graphs and also bridge the gap between vision and language by generating text to describe images details

Try Model

Fireworks Features

On-demand Deployment

On-demand deployments give you dedicated GPUs for Llama 3.2 11B Vision Instruct using Fireworks' reliable, high-performance system with no rate limits.

Learn More

Info

Provider

Meta

Model Type

LLMVision

Context Length

131072

Pricing Per 1M Tokens

$0.2