FFree
Glm-flashClosed Weights

GLM 4.1V Thinking Flash

glm-4.1v-thinking-flash

Best Input Price

$0.30 / 1M

Compact GPT model for low-latency assistance and high-volume workloads

Context Window

64,000 tokens

Reasoning

No

Tool Calling

No

Released

2025-07-09

Inference Providers (1)

ProviderModel IDContextInput / 1MOutput / 1MAction
NanoGPTglm-4.1v-thinking-flash64,000$0.30$0.30Docs ↗