Korean Telecommunications Domain LLM Question Collection Text Dataset
A Korean telecommunications domain question collection dataset built for LLM training, with domain-specific questions directly collected and curated by professional annotators.
Flitto Data Marketplace offers license-verified AI training and evaluation datasets in text, speech, image, and video across 100+ languages and 23+ domains. Try a free sample, then contact us to purchase.
A Korean telecommunications domain question collection dataset built for LLM training, with domain-specific questions directly collected and curated by professional annotators.
A Korean QA text dataset built from Korean news articles, with related FAQs generated and curated by professional annotators.
A legal document text dataset built by OCR-processing, hierarchically structuring, and translating EU privacy-related legal documents such as the GDPR and AI Act into Korean and English.
A low-resource language parallel corpus dataset built by translating Korean written and spoken-style source texts into eight low-resource languages, including Vietnamese, Indonesian, Thai, and Hindi.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Gulf Arabic.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Egyptian Arabic.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Bengali.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Hindi.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Indonesian.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Japanese.