Schematron V2 Turbo

active
Context window128K tokens
Input / 1M tokens$0.03
Output / 1M tokens$0.15

Version History

v2-turbominor

Inference.net released Schematron V2 Turbo, a 3B-parameter model specialized for high-volume HTML-to-JSON extraction with a 128K context window. Extraction rules must be defined via a JSON schema in response_format rather than through prompts.

Benchmark Scores

Full leaderboard →
7.0 tokens_per_sec
Speed (tok/s)

Coverage