Meta: Llama 3.2 11B Vision Instruct
Metallm
About
Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...
Try Meta: Llama 3.2 11B Vision Instruct
Test this model in the Sandbase Playground with your own prompts.
Open in Playground