FFree
VisparkClosed Weights

Vision Small

vispark/vision-small

Best Input Price

$1.05 / 1M

Fast, low-cost multimodal model for understanding text, images, audio, video, and PDFs, with tool calling and a 1M-token context window.

Context Window

1,000,000 tokens

Reasoning

Supported

Tool Calling

Supported

Released

2024-05-15

Inference Providers (1)

ProviderModel IDContextInput / 1MOutput / 1MAction
Visparkvispark/vision-small1,000,000$1.05$3.16Docs ↗