modelbenchmark.io

Schematron V2 Small

schematron-v2-small

Compare

Inference.net's 3B-parameter HTML-to-JSON extraction model, focused on accuracy for complex schemas and long web pages. It turns HTML into typed, structured data for web scraping and product catalog ingestion, with a 128K-token context window. Supply HTML in the user message and extraction instructions in a JSON schema via response_format; it does not follow ordinary chat or system prompts.

Specification

most-agreed values

Context
128K
Max output
4K
Released
2026-09-12
Knowledge cutoff
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.05
Output
$0.23
Cache read
$0.025

Available from 3 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
kilo · inference-net$0.05$0.23$0.05128K4K
nano-gpt · inference-net$0.05$0.23$0.025128K4K
openrouter · inference-net$0.05$0.23$0.05128K4K