paper-with-me

Papers

Evaluating Text Classification Robustness to Part-of-Speech Adversarial Examples

2024-08-15 · Anahita Samadi, Allison Sullivan

As machine learning systems become more widely used, especially for safety critical applications, there is a growing need to ensure that these systems behave as intended, even in the face of adversarial examples. Adversarial examples are inputs that are designed to trick the decision making process, and are intended to be imperceptible to humans. However, for text-based classification systems, changes to the input, a string of text, are always perceptible. Therefore, text-based adversarial examples instead focus on trying to preserve semantics. Unfortunately, recent work has shown this goal is often not met. To improve the quality of text-based adversarial examples, we need to know what elements of the input text are worth focusing on. To address this, in this paper, we explore what parts of speech have the highest impact of text-based classifiers. Our experiments highlight a distinct bias in CNN algorithms against certain parts of speech tokens within review datasets. This finding underscores a critical vulnerability in the linguistic processing capabilities of CNNs.

📄 PDF Abstract BibTeX arXiv:2408.08374

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Makingtext-classificationText Classification

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Evaluating Japanese Dialect Robustness Across Speech and Text-based Large Language Models

2026-06-24 · Tomoya Mizumoto, Yusuke Fujita, Hao Shi, Lianbo Liu 외 arxiv

Dialogue systems based on large language models (LLMs) have advanced significantly in recent years. However, dialectal variation remains a major challenge, particularly for systems that process spoken input. LLM-based sp…

VocalBench-DF: A Benchmark for Evaluating Speech LLM Robustness to Disfluency

2025-10-17 · Hongcheng Liu, Yixuan Hou, Heyang Liu, Yuhao Wang 외 arxiv

While Speech Large Language Models (Speech-LLMs) show strong performance in many applications, their robustness is critically under-tested, especially to speech disfluency. Existing evaluations often rely on idealized in…

Listening or Reading? Evaluating Speech Awareness in Chain-of-Thought Speech-to-Text Translation

2025-10-03 · Jacobo Romero-Díaz, Gerard I. Gállego, Oriol Pareras, Federico Costa 외 arxiv

Speech-to-Text Translation (S2TT) systems built from Automatic Speech Recognition (ASR) and Text-to-Text Translation (T2TT) modules face two major limitations: error propagation and the inability to exploit prosodic or o…

Speech-to-Text TranslationSpeech Recognition

ITALIC: An Italian Intent Classification Dataset

2023-06-14 · Alkis Koudounas, Moreno La Quatra, Lorenzo Vaiani, Luca Colomba 외

Recent large-scale Spoken Language Understanding datasets focus predominantly on English and do not account for language-specific phenomena such as particular phonemes or words in different lects. We introduce ITALIC, th…

Classificationintent-classificationIntent Classificationspeech-recognition+2

Evaluating the Effectiveness of Pre-Trained Audio Embeddings for Classification of Parkinson's Disease Speech Data

2025-06-02 · Emmy Postma, Cristian Tejedor-Garcia

Speech impairments are prevalent biomarkers for Parkinson's Disease (PD), motivating the development of diagnostic techniques using speech data for clinical applications. Although deep acoustic features have shown promis…

Diagnostic