paper-with-me

Papers Text Augmentation

“Text Augmentation” 태그가 달린 논문 97편 · 필터 해제

BrightCookies at SemEval-2025 Task 9: Exploring Data Augmentation for Food Hazard Classification

2025-04-29 · Foteini Papadopoulou, Osman Mutlu, Neris Özen, Bas H. M. van der Velden 외

This paper presents our system developed for the SemEval-2025 Task 9: The Food Hazard Detection Challenge. The shared task's objective is to evaluate explainable classification systems for classifying hazards and product…

Data AugmentationText Augmentation

Batch Aggregation: An Approach to Enhance Text Classification with Correlated Augmented Data

2025-04-07 · Charco Hui, Yalu Wen

Natural language processing models often face challenges due to limited labeled data, especially in domain specific areas, e.g., clinical trials. To overcome this, text augmentation techniques are commonly used to increa…

ClassificationText Augmentationtext-classificationText Classification

Toward General and Robust LLM-enhanced Text-attributed Graph Learning

2025-04-03 · Zihao Zhang, Xunkai Li, Rong-Hua Li, Bing Zhou 외

Recent advancements in Large Language Models (LLMs) and the proliferation of Text-Attributed Graphs (TAGs) across various domains have positioned LLM-enhanced TAG learning as a critical research area. By utilizing rich g…

Graph LearningTAGText Augmentation

Words or Vision: Do Vision-Language Models Have Blind Faith in Text?

2025-03-04 · CVPR 2025 1 · Ailin Deng, Tri Cao, Zhirui Chen, Bryan Hooi

Vision-Language Models (VLMs) excel in integrating visual and textual information for vision-centric tasks, but their handling of inconsistencies between modalities is underexplored. We investigate VLMs' modality prefere…

Language ModelingLanguage ModellingText Augmentation

Laser: Efficient Language-Guided Segmentation in Neural Radiance Fields

2025-01-31 · Xingyu Miao, Haoran Duan, Yang Bai, Tejal Shah 외

In this work, we propose a method that leverages CLIP feature distillation, achieving efficient 3D segmentation through language guidance. Unlike previous methods that rely on multi-scale CLIP features and are limited by…

SegmentationText Augmentation

Image, Text, and Speech Data Augmentation using Multimodal LLMs for Deep Learning: A Survey

2025-01-29 · Ranjan Sapkota, Shaina Raza, Maged Shoman, Achyut Paudel 외

In the past five years, research has shifted from traditional Machine Learning (ML) and Deep Learning (DL) approaches to leveraging Large Language Models (LLMs) , including multimodality, for data augmentation to enhance…

Data AugmentationImage AugmentationText Augmentation

Multimodal AI on Wound Images and Clinical Notes for Home Patient Referral

2025-01-22 · Reza Saadati Fard, Emmanuel Agu, Palawat Busaranuvong, Deepak Kumar 외

Chronic wounds affect 8.5 million Americans, particularly the elderly and patients with diabetes. These wounds can take up to nine months to heal, making regular care essential to ensure healing and prevent severe outcom…

Text AugmentationTransfer Learning

TARDiS : Text Augmentation for Refining Diversity and Separability

2025-01-06 · KyungMin Kim, SangHun Im, Gibaeg Kim, Heung-Seon Oh

Text augmentation (TA) is a critical technique for text classification, especially in few-shot settings. This paper introduces a novel LLM-based TA method, TARDiS, to address challenges inherent in the generation and ali…

DiversityFew-Shot Text ClassificationText Augmentationtext-classification+1

Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP

2024-12-13 · Yating Yu, Congqi Cao, Yueran Zhang, Qinyi Lv 외

Zero-shot action recognition (ZSAR) requires collaborative multi-modal spatiotemporal understanding. However, finetuning CLIP directly for ZSAR yields suboptimal performance, given its inherent constraints in capturing e…

Action RecognitionText AugmentationZero-Shot Action Recognition

An Experimental Study on Data Augmentation Techniques for Named Entity Recognition on Low-Resource Domains

2024-11-21 · Arthur Elwing Torres, Edleno Silva de Moura, Altigran Soares da Silva, Mario A. Nascimento 외

Named Entity Recognition (NER) is a machine learning task that traditionally relies on supervised learning and annotated data. Acquiring such data is often a challenge, particularly in specialized fields like medical, le…

Data Augmentationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Use Random Selection for Now: Investigation of Few-Shot Selection Strategies in LLM-based Text Augmentation for Classification

2024-10-14 · Jan Cegin, Branislav Pecher, Jakub Simko, Ivan Srba 외

The generative large language models (LLMs) are increasingly used for data augmentation tasks, where text samples are paraphrased (or generated anew) and then used for classifier fine-tuning. Existing works on augmentati…

Data AugmentationFew-Shot LearningText Augmentation

CleanerCLIP: Fine-grained Counterfactual Semantic Augmentation for Backdoor Defense in Contrastive Learning

2024-09-26 · Yuan Xun, Siyuan Liang, Xiaojun Jia, Xinwei Liu 외

Pre-trained large models for multimodal contrastive learning, such as CLIP, have been widely recognized in the industry as highly susceptible to data-poisoned backdoor attacks. This poses significant risks to downstream …

backdoor defenseContrastive LearningcounterfactualText Augmentation+2

Augment, Drop & Swap: Improving Diversity in LLM Captions for Efficient Music-Text Representation Learning

2024-09-17 · Ilaria Manco, Justin Salamon, Oriol Nieto

Audio-text contrastive models have become a powerful approach in music representation learning. Despite their empirical success, however, little is known about the influence of key design choices on the quality of music-…

DiversityRepresentation LearningText Augmentation

LLMs vs Established Text Augmentation Techniques for Classification: When do the Benefits Outweight the Costs?

2024-08-29 · Jan Cegin, Jakub Simko, Peter Brusilovsky

The generative large language models (LLMs) are increasingly being used for data augmentation tasks, where text samples are LLM-paraphrased and then used for classifier fine-tuning. However, a research that would confirm…

Data AugmentationText Augmentation

QAEA-DR: A Unified Text Augmentation Framework for Dense Retrieval

2024-07-29 · Hongming Tan, Shaoxiong Zhan, Hai Lin, Hai-Tao Zheng 외

In dense retrieval, embedding long texts into dense vectors can result in information loss, leading to inaccurate query-text matching. Additionally, low-quality texts with excessive noise or sparse key information are un…

Answer GenerationEvent ExtractionQuestion-Answer-GenerationRetrieval+5

Mitigating Data Imbalance for Software Vulnerability Assessment: Does Data Augmentation Help?

2024-07-15 · Triet H. M. Le, M. Ali Babar

Background: Software Vulnerability (SV) assessment is increasingly adopted to address the ever-increasing volume and complexity of SVs. Data-driven approaches have been widely used to automate SV assessment tasks, partic…

Data AugmentationText Augmentation

Performance Improvement of Language-Queried Audio Source Separation Based on Caption Augmentation From Large Language Models for DCASE Challenge 2024 Task 9

2024-06-17 · Do Hyun Lee, Yoonah Song, Hong Kook Kim

We present a prompt-engineering-based text-augmentation approach applied to a language-queried audio source separation (LASS) task. To enhance the performance of LASS, the proposed approach utilizes large language models…

Audio Source SeparationPrompt EngineeringSentenceText Augmentation

An efficient text augmentation approach for contextualized Mandarin speech recognition

2024-06-14 · Naijun Zheng, Xucheng Wan, Kai Liu, Ziqing Du 외

Although contextualized automatic speech recognition (ASR) systems are commonly used to improve the recognition of uncommon words, their effectiveness is hindered by the inherent limitations of speech-text data availabil…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1

ExplainableDetector: Exploring Transformer-based Language Modeling Approach for SMS Spam Detection with Explainability Analysis

2024-05-12 · Mohammad Amaz Uddin, Muhammad Nazrul Islam, Leandros Maglaras, Helge Janicke 외

SMS, or short messaging service, is a widely used and cost-effective communication medium that has sadly turned into a haven for unwanted messages, commonly known as SMS spam. With the rapid adoption of smartphones and I…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Language ModelingLanguage Modelling+2

Context-Aware Clustering using Large Language Models

2024-05-02 · Sindhu Tipirneni, Ravinarayana Adkathimar, Nurendra Choudhary, Gaurush Hiranandani 외

Despite the remarkable success of Large Language Models (LLMs) in text understanding and generation, their potential for text clustering tasks remains underexplored. We observed that powerful closed-source LLMs provide g…

ClusteringLanguage ModelingLanguage ModellingText Augmentation+2
1–20 / 97 다음 →