DasanCallDial: Large-Scale Korean ASR Error Correction Dataset
September 10, 2026
DasanCallDial introduces a benchmark of 1,974 dialogues and 115,460 utterances for text-based Korean ASR error correction. It addresses the lack of dialogue-level annotated corpora required to refine transcription in privacy-restricted environments like call centers.
HOW THIS AFFECTS YOU
●
builderYou can use this dataset to improve text-based post-editing for Korean transcription services.
●
researcherThis provides a specialized resource for studying low-resource language error correction in dialogue contexts.