Best Input Price
$0.02 / 1M
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts model with approximately 5.1B parameters active per token. It prioritizes token efficiency and production-scale agentic inference, helping coding and tool-using agents complete more work within constrained latency and serving budgets.
Context Window
262,144 tokens
Reasoning
No
Tool Calling
Supported
Released
2026-07-23
Inference Providers (4)
| Provider | Model ID | Context | Input / 1M | Output / 1M | Action |
|---|---|---|---|---|---|
| OpenRouter | inclusionai/ling-3.0-flash | 262,144 | $0.02 | $0.06 | Docs ↗ |
| Vercel AI Gateway | inclusionai/ling-3.0-flash | 256,000 | $0.02 | $0.06 | Docs ↗ |
| Kilo Gateway | inclusionai/ling-3.0-flash | 262,144 | $0.06 | $0.18 | Docs ↗ |
| NanoGPT | inclusionai/ling-3.0-flash | 262,144 | $0.07 | $0.22 | Docs ↗ |