FFree
InclusionAIClosed Weights

Ling 3.0 Flash

inclusionai/ling-3.0-flash

Best Input Price

$0.02 / 1M

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts model with approximately 5.1B parameters active per token. It prioritizes token efficiency and production-scale agentic inference, helping coding and tool-using agents complete more work within constrained latency and serving budgets.

Context Window

262,144 tokens

Reasoning

No

Tool Calling

Supported

Released

2026-07-23

Inference Providers (4)

ProviderModel IDContextInput / 1MOutput / 1MAction
OpenRouterinclusionai/ling-3.0-flash262,144$0.02$0.06Docs ↗
Vercel AI Gatewayinclusionai/ling-3.0-flash256,000$0.02$0.06Docs ↗
Kilo Gatewayinclusionai/ling-3.0-flash262,144$0.06$0.18Docs ↗
NanoGPTinclusionai/ling-3.0-flash262,144$0.07$0.22Docs ↗