paper-with-me

Papers

COMPILING: A Benchmark Dataset for Chinese Complexity Controllable Definition Generation

2022-09-29 · CCL 2022 10 · Jiaxin Yuan, Cunliang Kong, Chenhui Xie, Liner Yang, Erhong Yang

The definition generation task aims to generate a word's definition within a specific context automatically. However, owing to the lack of datasets for different complexities, the definitions produced by models tend to keep the same complexity level. This paper proposes a novel task of generating definitions for a word with controllable complexity levels. Correspondingly, we introduce COMPILING, a dataset given detailed information about Chinese definitions, and each definition is labeled with its complexity levels. The COMPILING dataset includes 74,303 words and 106,882 definitions. To the best of our knowledge, it is the largest dataset of the Chinese definition generation task. We select various representative generation methods as baselines for this task and conduct evaluations, which illustrates that our dataset plays an outstanding role in assisting models in generating different complexity-level definitions. We believe that the COMPILING dataset will benefit further research in complexity controllable definition generation.

📄 PDF Abstract BibTeX arXiv:2209.14614

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CCLAP: Controllable Chinese Landscape Painting Generation via Latent Diffusion Model

2023-04-09 · Zhongqi Wang, Jie Zhang, Zhilong Ji, Jinfeng Bai 외

With the development of deep generative models, recent years have seen great success of Chinese landscape painting generation. However, few works focus on controllable Chinese landscape painting generation due to the lac…

Chinese Landscape Painting Generation

Spontaneous Speech Corpora for language learners of Spanish, Chinese and Japanese

2012-05-01 · LREC 2012 5 · Moreno-S, Antonio oval, Leonardo Campillos Llanos, Yang Dong 외

This paper presents a method for designing, compiling and annotating corpora intended for language learners. In particular, we focus on spoken corpora for being used as complementary material in the classroom as well as …

HellaSwag-Pro: A Large-Scale Bilingual Benchmark for Evaluating the Robustness of LLMs in Commonsense Reasoning

2025-02-17 · Xiaoyuan Li, Moxin Li, Rui Men, Yichang Zhang 외

Large language models (LLMs) have shown remarkable capabilities in commonsense reasoning; however, some variations in questions can trigger incorrect responses. Do these models truly understand commonsense knowledge, or …

HellaSwag

HSKBenchmark: Modeling and Benchmarking Chinese Second Language Acquisition in Large Language Models through Curriculum Tuning

2025-11-19 · Qihao Yang, Xuelin Wang, Jiale Chen, Xuelian Dong 외 arxiv

Language acquisition is vital to revealing the nature of human language intelligence and has recently emerged as a promising perspective for improving the interpretability of large language models (LLMs). However, it is …

Language Acquisition

Efficient and practical quantum compiler towards multi-qubit systems with deep reinforcement learning

2022-04-14 · Qiuhao Chen, Yuxuan Du, Qi Zhao, Yuling Jiao 외

Efficient quantum compiling tactics greatly enhance the capability of quantum computers to execute complicated quantum algorithms. Due to its fundamental importance, a plethora of quantum compilers has been designed in p…

Deep Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning (RL)