Best Input Price
$0.00 / 1M
Open multimodal Llama model for image understanding, captioning, and visual QA
Context Window
131,072 tokens
Reasoning
No
Tool Calling
No
Released
2024-09-25
Inference Providers (4)
| Provider | Model ID | Context | Input / 1M | Output / 1M | Action |
|---|---|---|---|---|---|
| Nvidia | meta/llama-3.2-11b-vision-instruct | 128,000 | $0.00 | $0.00 | Docs ↗ |
| Cloudflare Workers AI | @cf/meta/llama-3.2-11b-vision-instruct | 128,000 | $0.05 | $0.68 | Docs ↗ |
| Inference | meta/llama-3.2-11b-vision-instruct | 16,000 | $0.06 | $0.06 | Docs ↗ |
| Eden AI | deepinfra/meta-llama/Llama-3.2-11B-Vision-Instruct | 131,072 | $0.34 | $0.34 | Docs ↗ |