paper-with-me

홈 › Papers

ORCHID: A Chinese Debate Corpus for Target-Independent Stance Detection and Argumentative Dialogue Summarization

2024-10-17 · Xiutian Zhao, Ke Wang, Wei Peng

Dialogue agents have been receiving increasing attention for years, and this trend has been further boosted by the recent progress of large language models (LLMs). Stance detection and dialogue summarization are two core tasks of dialogue agents in application scenarios that involve argumentative dialogues. However, research on these tasks is limited by the insufficiency of public datasets, especially for non-English languages. To address this language resource gap in Chinese, we present ORCHID (Oral Chinese Debate), the first Chinese dataset for benchmarking target-independent stance detection and debate summarization. Our dataset consists of 1,218 real-world debates that were conducted in Chinese on 476 unique topics, containing 2,436 stance-specific summaries and 14,133 fully annotated utterances. Besides providing a versatile testbed for future research, we also conduct an empirical study on the dataset and propose an integrated task. The results show the challenging nature of the dataset and suggest a potential of incorporating stance detection in summarization for argumentative dialogue.

📄 PDF Abstract BibTeX arXiv:2410.13667

Code (1)

xiutian/orchid 공식 구현

Tasks

BenchmarkingStance Detection

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Orchid2024: A cultivar-level dataset and methodology for fine-grained classification of Chinese Cymbidium Orchids

2024-08-13 · Plant Methods 2024 8 · Yingshu Peng, Yuxia Zhou, Li Zhang, Hongyan Fu 외

The authors dedicated over a year to collecting a cultivar image dataset for Chinese Cymbidium orchids named Orchid2024. This dataset contains over 150,000 images spanning 1,275 different categories, involving visits to …

Fine-Grained Image ClassificationImage Classificationparameter-efficient fine-tuning

R-Debater: Retrieval-Augmented Debate Generation through Argumentative Memory

2025-12-31 · Maoyuan Li, Zhongsheng Wang, Haoyuan Li, Jiamou Liu arxiv

We present R-Debater, an agentic framework for generating multi-turn debates built on argumentative memory. Grounded in rhetoric and memory studies, the system views debate as a process of recalling and adapting prior ar…

Constructing a Chinese---Japanese Parallel Corpus from Wikipedia

2014-05-01 · LREC 2014 5 · Chenhui Chu, Toshiaki Nakazawa, Sadao Kurohashi

Parallel corpora are crucial for statistical machine translation (SMT). However, they are quite scarce for most language pairs, such as Chinese―Japanese. As comparable corpora are far more available, many studies have be…

Machine TranslationSentenceTranslation

Target-based Sentiment Annotation in Chinese Financial News

2020-05-01 · LREC 2020 5 · Chaofa Yuan, Yu-Han Liu, Rongdi Yin, Jun Zhang 외

This paper presents the design and construction of a large-scale target-based sentiment annotation corpus on Chinese financial news text. Different from the most existing paragraph/document-based annotation corpus, in th…

Sentiment Analysis

Identification of Orchid Species Using Content-Based Flower Image Retrieval

2014-06-10 · D. H. Apriyanti, A. A. Arymurthy, L. T. Handoko

In this paper, we developed the system for recognizing the orchid species by using the images of flower. We used MSRM (Maximal Similarity based on Region Merging) method for segmenting the flower object from the backgrou…

Image RetrievalRetrieval