Inference.net: Schematron V2 Turbo

Text input Text output
Author's Description

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

Key Specifications
Cost
$$
Context
128K
Parameters
3B (Rumoured)
Released
Sep 11, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Temperature Stop Min P Response Format Top P Logit Bias Frequency Penalty Presence Penalty Max Tokens Structured Outputs Seed
Features

This model supports the following features:

Response Format Structured Outputs
Performance Summary

Inference.net's Schematron V2 Turbo, a 3B-parameter HTML-to-JSON extraction model, demonstrates exceptional operational efficiency. It consistently ranks among the fastest models and offers highly competitive pricing across all benchmarks, making it a strong contender for high-volume extraction workloads. The model also exhibits outstanding reliability with a 99% success rate, indicating minimal technical failures. However, its performance on general cognitive tasks is notably low. It struggles significantly with Instruction Following (0.0% accuracy) and Reasoning (14.0% accuracy). While its accuracy in Hallucinations (58.0%), General Knowledge (43.5%), Ethics (66.0%), Mathematics (49.0%), Email Classification (58.0%), and Coding (49.0%) is generally below average, its primary design as an HTML-to-JSON extraction model suggests these benchmarks may not reflect its core strength. Its key strength lies in its speed, cost-effectiveness, and reliability, which are crucial for its intended high-throughput extraction applications.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.03
Completion $0.15

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
InferenceNet
InferenceNet | inference-net/schematron-v2-turbo-20260902 128K $0.03 / 1M tokens $0.15 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by inference-net