Indonesian General-Domain Multi-Turn Chat Text Dataset
An Indonesian general-domain multi-turn chat text dataset built by professional annotators for daily life and general conversation domains.
Use verified, licensed data with confidence. You can download right away or check the data through inquiry.
An Indonesian general-domain multi-turn chat text dataset built by professional annotators for daily life and general conversation domains.
An Italian general-domain multi-turn chat text dataset built by professional annotators for daily life and general conversation domains.
A Japanese general-domain multi-turn chat text dataset built by professional annotators for daily life and general conversation domains.
A Spanish general-domain multi-turn chat text dataset built by professional annotators for daily life and general conversation domains.
A Vietnamese general-domain multi-turn chat text dataset built by professional annotators for daily life and general conversation domains.
A parallel translation corpus dataset built from Korean source texts and corresponding foreign-language translations across general domains, created by professional translators and native-language annotators.
A multilingual food and menu parallel translation corpus built by translating menu names extracted from sources such as AI Hub, the Korea Tourism Organization, and ImageLab into Korean, English, Chinese, and Japanese, with expert review by translators and native-language annotators.
A Korean healthcare and bio domain question collection dataset built for LLM training, with questions directly collected and curated by medical and biotechnology majors and industry professionals.
A Korean construction and engineering domain question collection dataset built for LLM training, with questions directly collected and curated by relevant majors and industry professionals.
A Korean education domain question collection dataset built for LLM training, with domain-specific questions directly collected and curated by professional annotators.