Kilo Gateway
Inference.net: Schematron V2 Turbo
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
inference-net/schematron-v2-turboContext
128K
Max output
8.2K
Input $/M
$0.03
Output $/M
$0.15
Cache read $/M
$0.03
Cache write $/M
—
Against the catalogue
Context window128K
Larger than 18% of the 7,784 models listed
Input price$0.03
Cheaper than 90% of priced models
Capabilities
Reasoningnot supported
Tool callingnot supported
Structured outputsupported
Attachmentsnot supported
Temperaturesupported
Open weightsnot supported
Interleaved thinkingnot supported
Modalities
Input
Text
Output
Text
Pricing detail
| Rate | Base $/M |
|---|---|
| Input | $0.03 |
| Output | $0.15 |
| Cache read | $0.03 |