Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Schematron V2 Turbo

Schematron V2 Turbo is a compact 3-billion-parameter model built specifically for converting raw HTML into structured JSON output, a narrow but high-demand task in web scraping, content migration, and document automation pipelines. Its design favors throughput over depth, making it suited for batch processing of many pages rather than single, deeply reasoned extractions. Extraction behavior is controlled by passing a JSON schema through the response format parameter, which steers the model toward consistent field layouts across runs and helps standardize output across large jobs.

The practical sweet spot for Schematron V2 Turbo is organizations that ingest large volumes of HTML and need predictable, schema-conformant JSON at scale without the latency cost of larger general-purpose models. Its small footprint and throughput-oriented positioning make it a good fit for server-side pipelines that process scraped pages, product catalogs, or templated documents where speed and cost dominate over conversational nuance. Compared with sibling offerings in the same family aimed at complex schemas and long pages, the Turbo variant trades some quality headroom for faster extraction on more uniform inputs.

OpenRouterinference-net/schematron-v2-turbo

Quick Info

Powered by
Provider
OpenRouter
Model key
inference-net/schematron-v2-turbo
Release date
Sep 12, 2026
Last updated
Sep 12, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.03
Output token cost
$0.15

Limits

Output tokens
8,192 tokens
Context window
128,000 tokens

Latest news about Schematron V2 Turbo

Vercel AI Gateway

CoverageBenchmark

LLMBase's third-party catalog profile explicitly names Inference.net's Schematron V2 Turbo under the exact slug "inference-net/schematron-v2-turbo," confirming it as a 3B-parameter text-only model purpose-built for HTML-to-JSON extraction that requires extraction instructions to be supplied via a JSON schema in the res The page notes that Schematron V2 Turbo is not currently available in LLMBase's own chat or inference API, positioning this entry as a research/comparison data point rather than a primary Inference.net announcement. As an aggregator restatement of baseline specifications rather than breaking news, the profile is useful

Vercel AI Gateway

CoverageBenchmark

Infron's performance/pricing page explicitly names Inference.net's Schematron V2 Turbo under the slug "inference-net/schematron-v2-turbo" and lists it as a text-generation, streaming, tool-calling, JSON-mode capable model with a single official service tier priced at $0.03 per million input tokens, $0.15 per million ou The page presents uptime (100.00% over 72 hours after routing/fallbacks) and provider details for Inference.net alongside links to explore a sibling model, Schematron V2 Small, which is also attributed to Inference.net. This is third-party routing/serving-layer data rather than a primary Inference.net announcement, and

OpenRouter

CoverageRelease Notes

A timeline of AI model releases lists "Inference.net: Schematron V2 Turbo" as the top entry, noting an API launch on 2026-09-12 with a 128k context window. The page is auto-refreshed from the OpenRouter models API and presents a chronological feed of new model additions, providing a concrete launch date and context len No benchmarks, pricing, architectural details, or capability claims for Schematron V2 Turbo are provided on the page, and no first-party Inference.net or OpenRouter announcement is linked. A sibling entry for "Inference.net: Schematron V2 Small" (also 128k context, same date) appears as entry #02, situating the Turbo v

Videos about Schematron V2 Turbo