VLM Run Gateway: Run GLM-OCR, DeepSeek-OCR-2, dots.mocr with an OpenAI Compatible API
Blog post from Hugging Face
VLM Run Gateway is presented as an OpenAI-compatible service for running open-weight OCR and vision-language models, including GLM-OCR, DeepSeek-OCR-2, dots.mocr, PaddleOCR-VL, and PP-OCRv6, through a single API. It aims to reduce the cost and operational complexity of document parsing by handling PDF rasterization, page distribution, ordering, streaming, retries, and memory failures, while allowing teams to switch models without rebuilding their workflows. The service supports structured JSON outputs for API integrations and an MCP server that enables compatible AI agents to read documents. Its authors argue that open-weight models can often handle OCR, extraction, layout analysis, and parsing at substantially lower cost than frontier VLM APIs, though frontier models remain preferable for more open-ended reasoning tasks. Because performance varies by document type, language, scan quality, and layout complexity, users are encouraged to evaluate multiple models on their own production documents, with cost data included in each response and a public accuracy leaderboard planned.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| MCP | 6 | 8,107 | 809 | 199 | -26% |
| Vector Search | 1 | 2,312 | 357 | 123 | +3% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.