Korean Telecommunications Domain LLM Question Collection Text Dataset
A Korean telecommunications domain question collection dataset built for LLM training, with domain-specific questions directly collected and curated by professional annotators.
Use verified, licensed data with confidence. You can download right away or check the data through inquiry.
A Korean telecommunications domain question collection dataset built for LLM training, with domain-specific questions directly collected and curated by professional annotators.
A Korean QA text dataset built from Korean news articles, with related FAQs generated and curated by professional annotators.
A legal document text dataset built by OCR-processing, hierarchically structuring, and translating EU privacy-related legal documents such as the GDPR and AI Act into Korean and English.
A low-resource language parallel corpus dataset built by translating Korean written and spoken-style source texts into eight low-resource languages, including Vietnamese, Indonesian, Thai, and Hindi.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Gulf Arabic.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Egyptian Arabic.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Bengali.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Hindi.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Indonesian.
A high-difficulty text dataset developed to train expert-level reasoning in LLMs based on doctoral examination questions and solutions in Japanese.