Inference.net: Schematron V2 Small

Text input Text output
Author's Description

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...

Key Specifications
Cost
$$
Context
128K
Parameters
3B (Rumoured)
Released
Sep 11, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Temperature Stop Min P Response Format Top P Logit Bias Frequency Penalty Presence Penalty Max Tokens Structured Outputs Seed
Features

This model supports the following features:

Response Format Structured Outputs
Performance Summary

Inference.net's Schematron V2 Small, a 3B-parameter HTML-to-JSON extraction model, demonstrates exceptional operational efficiency. It consistently ranks among the fastest models, achieving the 93rd percentile across 8 benchmarks, and offers highly competitive pricing, placing in the 88th percentile. The model exhibits outstanding reliability with a 100% success rate across all benchmarks, indicating minimal technical failures. While optimized for HTML-to-JSON extraction, its performance across general benchmarks reveals areas for improvement. The model struggles significantly with Instruction Following (1% accuracy), Reasoning (16% accuracy), and Hallucinations (32% accuracy), suggesting limitations in complex cognitive tasks and uncertainty handling. Performance in General Knowledge (82%), Ethics (85%), and Mathematics (38%) is below average compared to other models. Its Email Classification (59%) and Coding (46%) capabilities also show room for growth. Schematron V2 Small's core strength lies in its intended purpose of extraction quality for complex schemas and long pages, which is not directly measured here, alongside its impressive speed, cost-effectiveness, and reliability.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.05
Completion $0.23

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
InferenceNet
InferenceNet | inference-net/schematron-v2-small-20260902 128K $0.05 / 1M tokens $0.23 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by inference-net