meta-llama

Meta: Llama 3.2 11B Vision Instruct

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

Input Cost

$0.25

per 1M tokens

Output Cost

$0.25

per 1M tokens

Context Window

131,072

tokens

Compare vs GPT-4o

                Developer ID: meta-llama/llama-3.2-11b-vision-instruct            

Related Models

meta-llama

$0.48/1M

Llama Guard 3 8B

Llama Guard 3 is a Llama-3.1-8B pretrained model, fine-tuned for content safety classifica...

📝 131,072 ctx Compare →

meta-llama

$0.12/1M

Meta: Llama 3.3 70B Instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction...

📝 131,072 ctx Compare →

meta-llama

$0.08/1M

Meta: Llama 4 Scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by...

📝 327,680 ctx Compare →