Inference.net: Schematron V2 Small
by Inference Nettext->text1 endpoint
Overview
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema in response_format rather than through system or user prompts.
Capabilities
Text generation
Structured output
Prompt caching
Long context
Top-P sampling
Stop sequences
Deterministic seed
Modalities
Input
Text
Output
Text
Technical Specifications
Context Window
128K tokens
Max Output
4.1K tokens
Tokenizer
Other
Uptime (24h)
10000.0%
Supported Parameters
frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyresponse_formatseedstopstructured_outputstemperaturetop_ktop_p
Pricing
Per 1M tokens
Input$0.05
Output$0.23
Cache read$0.05
Added
September 12, 2026
Last synced 9/18/2026