paper-with-me

홈 › Papers

Building Korean Linguistic Resource for NLU Data Generation of Banking App CS Dialog System

2022-10-01 · PANDL (COLING) 2022 10 · Jeongwoo Yoon, Onyu Park, Changhoe Hwang, Gwanghoon Yoo, Eric Laporte, Jeesun Nam

Natural language understanding (NLU) is integral to task-oriented dialog systems, but demands a considerable amount of annotated training data to increase the coverage of diverse utterances. In this study, we report the construction of a linguistic resource named FIAD (Financial Annotated Dataset) and its use to generate a Korean annotated training data for NLU in the banking customer service (CS) domain. By an empirical examination of a corpus of banking app reviews, we identified three linguistic patterns occurring in Korean request utterances: TOPIC (ENTITY, FEATURE), EVENT, and DISCOURSE MARKER. We represented them in LGGs (Local Grammar Graphs) to generate annotated data covering diverse intents and entities. To assess the practicality of the resource, we evaluate the performances of DIET-only (Intent: 0.91 /Topic [entity+feature]: 0.83), DIET+ HANBERT (I:0.94/T:0.85), DIET+ KoBERT (I:0.94/T:0.86), and DIET+ KorBERT (I:0.95/T:0.84) models trained on FIAD-generated data to extract various types of semantic items.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Natural Language Understanding

Methods 이 논문이 사용한 방법론

Golden Queue Managers 설명 없음

Similar Papers 제목 키워드 기반

Building Korean linguistic resource for NLU data generation of banking app CS dialog system

2026-05-11 · Jeongwoo Yoon, On-yu Park, Changhoe Hwang, Gwanghoon Yoo 외 arxiv

Natural language understanding (NLU) is integral to task-oriented dialog systems, but demands a considerable amount of annotated training data to increase the coverage of diverse utterances. In this study, we report the …

Natural Language Understanding

A Dog Is Passing Over The Jet? A Text-Generation Dataset for Korean Commonsense Reasoning and Evaluation

2022-07-01 · Findings (NAACL) 2022 7 · Jaehyung Seo, Seounghoon Lee, Chanjun Park, Yoonna Jang 외

Recent natural language understanding (NLU) research on the Korean language has been vigorously maturing with the advancements of pretrained language models and datasets. However, Korean pretrained language models still …

Language Model EvaluationLanguage ModelingLanguage ModellingNatural Language Understanding+2

Korean Language Modeling via Syntactic Guide

2022-06-01 · LREC 2022 6 · Hyeondey Kim, Seonhoon Kim, Inho Kang, Nojun Kwak 외

While pre-trained language models play a vital role in modern language processing tasks, but not every language can benefit from them. Most existing research on pre-trained language models focuses primarily on widely-use…

Language ModelingLanguage ModellingPOS

DECO-MWE: building a linguistic resource of Korean multiword expressions for feature-based sentiment analysis

2026-05-11 · Jaeho Han, Changhoe Hwang, Seongyong Choi, Gwanghoon Yoo 외 arxiv

This paper aims to construct a linguistic resource of Korean Multiword Expressions for Feature-Based Sentiment Analysis (FBSA): DECO-MWE. Dealing with multiword expressions (MWEs) has been a critical issue in FBSA since …

Sentiment Analysis

Developing Language Resources and NLP Tools for the North Korean Language

2022-06-01 · LREC 2022 6 · Arda Akdemir, Yeojoo Jeon, Tetsuo Shibuya

Since the division of Korea, the two Korean languages have diverged significantly over the last 70 years. However, due to the lack of linguistic source of the North Korean language, there is no DPRK-based language model.…

Language ModelingLanguage ModellingMasked Language ModelingSentiment Analysis