Meta: Llama 3.2 90B Vision Instruct

128K Context
0.35/M Input Tokens
0.4/M Output Tokens
0.506/K Image Tokens

Model Unavailable

The Llama 90B Vision model is a top-tier, 90-billion-parameter multimodal model designed for the most challenging visual reasoning and language tasks. It offers unparalleled accuracy in image captioning, visual question answering, and advanced image-text comprehension. Pre-trained on vast multimodal datasets and fine-tuned with human feedback, the Llama 90B Vision is engineered to handle the most demanding image-based AI tasks.

This model is perfect for industries requiring cutting-edge multimodal AI capabilities, particularly those dealing with complex, real-time visual and textual analysis.

Click here for the original model card.

Usage of this model is subject to Meta’s Acceptable Use Policy.

Meta: Llama 3.1 70B Instruct

Text 2 text

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrate ...

Meta llama 128K context $0.3/M input tokens $0.3/M output tokens

Meta: Llama 3.1 8B Instruct

Text 2 text

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compar ...

Meta llama 128K context $0.055/M input tokens $0.055/M output tokens

Meta: Llama 3.2 11B Vision Instruct

Text image 2 text

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and visual ...

Meta llama 128K context $0.055/M input tokens $0.055/M output tokens $0.079/K image tokens

Meta: Llama 3.2 1B Instruct

Text 2 text

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its small ...

Meta llama 128K context $0.01/M input tokens $0.02/M output tokens

Meta: Llama 3.2 3B Instruct

Text 2 text

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. ...

Meta llama 128K context $0.03/M input tokens $0.05/M output tokens

Meta: Llama 3.2 90B Vision Instruct

Tags :

Share :

Related Posts

Meta: Llama 3.1 70B Instruct

Meta: Llama 3.1 8B Instruct

Meta: Llama 3.2 11B Vision Instruct

Meta: Llama 3.2 1B Instruct

Meta: Llama 3.2 3B Instruct