Home / Companies / Inference / Blog / Post Details
Content Deep Dive

Schematron V2: Frontier HTML-to-JSON extraction at a fraction of the cost

Blog post from Inference

Post Details
Company
Date Published
Author
Amar Singh
Word Count
1,041
Company Posts That Month
2
Language
English
Hacker News Points
-
Post removed?
No
Summary

Schematron V2 is the latest iteration of specialized HTML-to-JSON extraction models, offering enhanced performance with two new variants, Schematron V2 Small and Schematron V2 Turbo, which significantly improve upon the previous generation's speed and quality. Optimized for cost and latency, these models maintain high extraction quality while providing faster processing capabilities, making them suitable for large-scale data extraction tasks. Schematron V2 Small nearly matches the quality of the original 8B model with faster performance, while Schematron V2 Turbo focuses on maximizing throughput, achieving 4.14 requests per second, which is 2.5 times faster than its predecessor. Both models are available through a serverless API on Inference.net, with pricing designed to make web-scale extraction more accessible. Future developments include Schematron Pro, which aims to offer even higher accuracy without sacrificing throughput.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 8 5,932 1,046 223 -2%
Serverless 3 678 211 91 -7%
AI Model Fine-tuning 1 420 130 55 -54%
Real-time 1 6,296 1,346 246 -2%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.