Best Input Price
$0.05 / 1M
Inference.net's 3B-parameter HTML-to-JSON extraction model, focused on accuracy for complex schemas and long web pages. It turns HTML into typed, structured data for web scraping and product catalog ingestion, with a 128K-token context window. Supply HTML in the user message and extraction instructions in a JSON schema via response_format; it does not follow ordinary chat or system prompts.
Context Window
128,000 tokens
Reasoning
No
Tool Calling
No
Released
2026-09-12
Inference Providers (4)
| Provider | Model ID | Context | Input / 1M | Output / 1M | Action |
|---|---|---|---|---|---|
| NanoGPT | inference-net/schematron-v2-small | 128,000 | $0.05 | $0.23 | Docs ↗ |
| Kilo Gateway | inference-net/schematron-v2-small | 128,000 | $0.05 | $0.23 | Docs ↗ |
| OpenRouter | inference-net/schematron-v2-small | 128,000 | $0.05 | $0.23 | Docs ↗ |
| Vercel AI Gateway | inference-net/schematron-v2-small | 128,000 | $0.05 | $0.23 | Docs ↗ |