Schematron V2: Frontier HTML-to-JSON extraction at a fraction of the cost
Blog post from Inference
Schematron V2 is the latest iteration of specialized HTML-to-JSON extraction models, offering enhanced performance with two new variants, Schematron V2 Small and Schematron V2 Turbo, which significantly improve upon the previous generation's speed and quality. Optimized for cost and latency, these models maintain high extraction quality while providing faster processing capabilities, making them suitable for large-scale data extraction tasks. Schematron V2 Small nearly matches the quality of the original 8B model with faster performance, while Schematron V2 Turbo focuses on maximizing throughput, achieving 4.14 requests per second, which is 2.5 times faster than its predecessor. Both models are available through a serverless API on Inference.net, with pricing designed to make web-scale extraction more accessible. Future developments include Schematron Pro, which aims to offer even higher accuracy without sacrificing throughput.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 8 | 5,932 | 1,046 | 223 | -2% |
| Serverless | 3 | 678 | 211 | 91 | -7% |
| AI Model Fine-tuning | 1 | 420 | 130 | 55 | -54% |
| Real-time | 1 | 6,296 | 1,346 | 246 | -2% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.