Home / Companies / Box / Blog / Post Details
Content Deep Dive

OCR-powered metadata extraction with Box AI and MCP

Blog post from Box

Post Details
Company
Box
Date Published
Author
Rui Barbosa, Senior Developer Advocate at Box
Word Count
815
Company Posts That Month
26
Language
English
Hacker News Points
-
Post removed?
No
Summary

Box AI has introduced OCR support for image files in its structured metadata extraction endpoints, enhancing the capabilities of its API by allowing direct processing of TIFF, PNG, and JPEG files without the need for conversion to PDF. This update supports multiple languages, including English, Japanese, Chinese, Korean, and Cyrillic scripts, and is available in both standard and enhanced structured extraction endpoints, although it is not part of the freeform extraction API. This development streamlines the workflow for developers by eliminating a conversion step, enabling the direct extraction of structured metadata from image-based documents like receipts, invoices, and forms. Box AI can now suggest metadata schemas for image files, facilitating the creation of templates that can be used to extract data with improved accuracy, especially when using sophisticated models for complex layouts. The integration also supports multilingual document processing, broadening the applicability of Box AI's document management systems across various markets.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
MCP 7 5,085 420 153 -2%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.