Inference.net: Schematron V2 Small

by Inference Nettext->text1 endpoint

Overview

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema in response_format rather than through system or user prompts.

Capabilities

Text generation
Structured output
Prompt caching
Long context
Top-P sampling
Stop sequences
Deterministic seed

Modalities

Input
Text
Output
Text

Technical Specifications

Context Window
128K tokens
Max Output
4.1K tokens
Tokenizer
Other
Uptime (24h)
10000.0%

Supported Parameters

frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyresponse_formatseedstopstructured_outputstemperaturetop_ktop_p

Pricing

Per 1M tokens
Input$0.05
Output$0.23
Cache read$0.05
Added
September 12, 2026
Last synced 9/18/2026