Data Marketplace

Use verified, licensed data with confidence. You can download right away or check the data through inquiry.

Sell your data on Flitto

Simply register your data and check its sale eligibility.

A total of 326 datasets
  • Pre-training DataImage

    Japanese Real-World OCR Image Dataset

    A Japanese real-world OCR image dataset built with OCR annotations for Japanese text detection, recognition, and Document AI model development.

  • Pre-training DataImage

    Japanese Food Image OCR Dataset

    A Japanese OCR image dataset built from collected images related to Japanese food, supporting visual text recognition and domain-specific image understanding.

  • Pre-training DataImage

    Japanese Landscape Image OCR Dataset

    A Japanese OCR image dataset built from collected images related to Japanese landscapes, supporting visual text recognition and domain-specific image understanding.

  • Pre-training DataImage

    Japanese Storefront Sign OCR Image Dataset

    A Japanese OCR image dataset built from collected real-world storefront sign images, supporting sign text recognition and scene text understanding.

  • Pre-training DataImage

    Korean Handwritten Text OCR Image Dataset

    A Korean handwritten text OCR image dataset built from real-world images with JSON annotations containing bounding boxes, language metadata, and reading direction information.

  • Pre-training DataImage

    Korean Document Table, Formula, and Graph OCR Image Dataset

    A Korean document OCR image dataset containing tables, LaTeX formulas, text, and graphs, built with JSON annotations containing bounding-box coordinates.

  • Pre-training DataImage

    Korean Menu OCR Image Dataset

    A Korean OCR image dataset for real-world menu images in the food and restaurant domain, reviewed and annotated by professional annotators.

  • Pre-training DataImage

    Korean Department Store Menu OCR Image Dataset

    A Korean OCR image dataset for real-world department store menu images in the food and restaurant domain, reviewed and annotated by professional annotators.

  • Alignment DataImage

    Korean Society and Culture Image Captioning Dataset

    A Korean image captioning dataset built with Korean captions for images related to Korean society, culture, and everyday visual contexts.

  • Pre-training DataImage

    Korean Handwriting Bounding Box Annotation OCR Image Dataset

    An OCR image dataset built by applying bounding box annotations to Korean handwriting images for multilingual document understanding.