Best Input Price
$0.03 / 1M
Inference.net's 3B-parameter HTML-to-JSON extraction model, optimized for throughput and low cost on high-volume workloads. It turns HTML into typed, structured data for web scraping and product catalog ingestion, with a 128K-token context window. Supply HTML in the user message and extraction instructions in a JSON schema via response_format; it does not follow ordinary chat or system prompts.
Context Window
128,000 tokens
Reasoning
No
Tool Calling
No
Released
2026-09-12
Inference Providers (4)
| Provider | Model ID | Context | Input / 1M | Output / 1M | Action |
|---|---|---|---|---|---|
| NanoGPT | inference-net/schematron-v2-turbo | 128,000 | $0.03 | $0.15 | Docs ↗ |
| Kilo Gateway | inference-net/schematron-v2-turbo | 128,000 | $0.03 | $0.15 | Docs ↗ |
| OpenRouter | inference-net/schematron-v2-turbo | 128,000 | $0.03 | $0.15 | Docs ↗ |
| Vercel AI Gateway | inference-net/schematron-v2-turbo | 128,000 | $0.03 | $0.15 | Docs ↗ |