Meta: Llama 3.2 90B Vision Instruct

meta-llama/llama-3.2-90b-vision-instruct

Created Sep 25, 202432,768 context

$0.35/M input tokens$0.40/M output tokens

The Llama 90B Vision model is a top-tier, 90-billion-parameter multimodal model designed for the most challenging visual reasoning and language tasks. It offers unparalleled accuracy in image captioning, visual question answering, and advanced image-text comprehension. Pre-trained on vast multimodal datasets and fine-tuned with human feedback, the Llama 90B Vision is engineered to handle the most demanding image-based AI tasks.

This model is perfect for industries requiring cutting-edge multimodal AI capabilities, particularly those dealing with complex, real-time visual and textual analysis.

Click here for the original model card.

Usage of this model is subject to Meta's Acceptable Use Policy.

Meta: Llama 3.2 90B Vision Instruct

meta-llama/llama-3.2-90b-vision-instruct

Created Sep 25, 202432,768 context

$0.35/M input tokens$0.40/M output tokens

This model is perfect for industries requiring cutting-edge multimodal AI capabilities, particularly those dealing with complex, real-time visual and textual analysis.

Click here for the original model card.

Usage of this model is subject to Meta's Acceptable Use Policy.

Meta: Llama 3.2 90B Vision Instruct

meta-llama/llama-3.2-90b-vision-instruct

Meta: Llama 3.2 90B Vision Instruct

meta-llama/llama-3.2-90b-vision-instruct

Recent activity on Llama 3.2 90B Vision Instruct

Total usage per day on OpenRouter

Recent activity on Llama 3.2 90B Vision Instruct

Total usage per day on OpenRouter