English Manufacturing Document OCR and Parsing Image Dataset
An English OCR and document parsing sample dataset for manufacturing documents, built with bounding-box JSON annotations for PoC, Document AI, and model evaluation use cases.
Flitto Data Marketplace offers license-verified AI training and evaluation datasets in text, speech, image, and video across 100+ languages and 23+ domains. Try a free sample, then contact us to purchase.
An English OCR and document parsing sample dataset for manufacturing documents, built with bounding-box JSON annotations for PoC, Document AI, and model evaluation use cases.
A Japanese handwritten document OCR image dataset with bounding-box annotations for multilingual document understanding, OCR model training, and layout-aware text recognition.
A Korean document summarization text dataset built by converting Korean news articles into markdown-based outputs such as key trends, summary reports, and document summaries.