Data Marketplace

Use verified, licensed data with confidence. You can download right away or check the data through inquiry.

Sell your data on Flitto

Simply register your data and check its sale eligibility.

A total of 323 datasets
  • Pre-training DataImage

    Japanese Storefront Sign OCR Image Dataset

    A Japanese OCR image dataset built from collected real-world storefront sign images, supporting sign text recognition and scene text understanding.

  • Pre-training DataImage

    Korean Handwritten Text OCR Image Dataset

    A Korean handwritten text OCR image dataset built from real-world images with JSON annotations containing bounding boxes, language metadata, and reading direction information.

  • Pre-training DataImage

    Korean Document Table, Formula, and Graph OCR Image Dataset

    A Korean document OCR image dataset containing tables, LaTeX formulas, text, and graphs, built with JSON annotations containing bounding-box coordinates.

  • Pre-training DataImage

    Korean Menu OCR Image Dataset

    A Korean OCR image dataset for real-world menu images in the food and restaurant domain, reviewed and annotated by professional annotators.

  • Pre-training DataImage

    Korean Department Store Menu OCR Image Dataset

    A Korean OCR image dataset for real-world department store menu images in the food and restaurant domain, reviewed and annotated by professional annotators.

  • Alignment DataImage

    Korean Society and Culture Image Captioning Dataset

    A Korean image captioning dataset built with Korean captions for images related to Korean society, culture, and everyday visual contexts.

  • Pre-training DataImage

    Korean Handwriting Bounding Box Annotation OCR Image Dataset

    An OCR image dataset built by applying bounding box annotations to Korean handwriting images for multilingual document understanding.

  • Pre-training DataImage

    Korean Handwriting Layout Annotation OCR Image Dataset

    A Korean handwritten document OCR image dataset with layout analysis annotations for multilingual document understanding, OCR model training, and document structure recognition.

  • Pre-training DataImage

    Korean Newspaper Bounding Box OCR Image Dataset

    A Korean newspaper OCR image dataset with bounding-box annotations for multilingual document understanding, text detection, and document layout-aware OCR.

  • Pre-training DataImage

    Korean Newspaper Layout OCR Image Dataset

    A Korean newspaper OCR image dataset with layout analysis annotations for multilingual document understanding, document structure recognition, and OCR model evaluation.