meta-llama

Meta: Llama 3.2 11B Vision Instruct

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

Input Cost
$0.25
per 1M tokens
Output Cost
$0.25
per 1M tokens
Context Window
131,072
tokens
Compare vs GPT-4o
Developer ID: meta-llama/llama-3.2-11b-vision-instruct

Related Models

meta-llama
$0.18/1M

Meta: Llama Guard 4 12B

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for conte...

📝 163,840 ctx Compare →
meta-llama
$0.03/1M

Meta: Llama 3.2 1B Instruct

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing nat...

📝 131,072 ctx Compare →
meta-llama
$0.51/1M

Meta: Llama 3 70B Instruct

Meta's latest class of model (Llama 3) launched with a variety of sizes & flavors. This 70...

📝 8,192 ctx Compare →