English Manufacturing Document OCR and Parsing Image Dataset
An English OCR and document parsing sample dataset for manufacturing documents, built with bounding-box JSON annotations for PoC, Document AI, and model evaluation use cases.
Use verified, licensed data with confidence. You can download right away or check the data through inquiry.
An English OCR and document parsing sample dataset for manufacturing documents, built with bounding-box JSON annotations for PoC, Document AI, and model evaluation use cases.
A Japanese handwritten document OCR image dataset with bounding-box annotations for multilingual document understanding, OCR model training, and layout-aware text recognition.
A Korean document summarization text dataset built by converting Korean news articles into markdown-based outputs such as key trends, summary reports, and document summaries.