All AI models · LLM benchmarks · Methodology

Schematron V2 Turbo

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

Model facts

Context window
128K tokens
Maximum output
8K tokens
Cheapest paid input
$0.03 per 1M tokens
Cheapest paid output
$0.15 per 1M tokens
Open weights
yes
Providers
1

Capabilities and modalities

structured output, temperature control, open weights, text.

Observed price history

  • 2026-09-12: $0.03 input / $0.15 output per 1M tokens

API providers

  • OpenRouter — model id inference-net/schematron-v2-turbo; $0.03 input / $0.15 output per 1M tokens; provider documentation