paper-with-me

Papers

Evaluation Metrics for Text Data Augmentation in NLP

2024-02-09 · Marcellus Amadeus, William Alberto Cruz Castañeda

Recent surveys on data augmentation for natural language processing have reported different techniques and advancements in the field. Several frameworks, tools, and repositories promote the implementation of text data augmentation pipelines. However, a lack of evaluation criteria and standards for method comparison due to different tasks, metrics, datasets, architectures, and experimental settings makes comparisons meaningless. Also, a lack of methods unification exists and text data augmentation research would benefit from unified metrics to compare different augmentation methods. Thus, academics and the industry endeavor relevant evaluation metrics for text data augmentation techniques. The contribution of this work is to provide a taxonomy of evaluation metrics for text augmentation methods and serve as a direction for a unified benchmark. The proposed taxonomy organizes categories that include tools for implementation and metrics calculation. Finally, with this study, we intend to present opportunities to explore the unification and standardization of text data augmentation metrics.

📄 PDF Abstract BibTeX arXiv:2402.06766

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationText Augmentation

Similar Papers 제목 키워드 기반

Advances in Diffusion Models for Image Data Augmentation: A Review of Methods, Models, Evaluation Metrics and Future Research Directions

2024-07-04 · Panagiotis Alimisis, Ioannis Mademlis, Panagiotis Radoglou-Grammatikis, Panagiotis Sarigiannidis 외

Image data augmentation constitutes a critical methodology in modern computer vision tasks, since it can facilitate towards enhancing the diversity and quality of training datasets; thereby, improving the performance and…

Data AugmentationDiversityImage Augmentation

Reducing Gender Bias in Word-Level Language Models with a Gender-Equalizing Loss Function

2019-05-30 · ACL 2019 7 · Yusu Qian, Urwa Muaz, Ben Zhang, Jae Won Hyun

Gender bias exists in natural language datasets which neural language models tend to learn, resulting in biased text generation. In this research, we propose a debiasing approach based on the loss function modification. …

Data AugmentationText Generation

Do Generative Metrics Predict YOLO Performance? An Evaluation Across Models, Augmentation Ratios, and Dataset Complexity

2026-02-20 · Vasile Marian, Yong-Bin Kang, Alexander Buddery arxiv

Synthetic images are increasingly used to augment object-detection training sets, but reliably evaluating a synthetic dataset before training remains difficult: standard global generative metrics (e.g., FID) often do not…

SDA: Improving Text Generation with Self Data Augmentation

2021-01-02 · Ping Yu, Ruiyi Zhang, Yang Zhao, Yizhe Zhang 외

Data augmentation has been widely used to improve deep neural networks in many research fields, such as computer vision. However, less work has been done in the context of text, partially due to its discrete nature and t…

Data AugmentationImitation LearningSentenceText Generation

Investigating Personalization Methods in Text to Music Generation

2023-09-20 · Manos Plitsis, Theodoros Kouzelis, Georgios Paraskevopoulos, Vassilis Katsouros 외

In this work, we investigate the personalization of text-to-music diffusion models in a few-shot setting. Motivated by recent advances in the computer vision domain, we are the first to explore the combination of pre-tra…

Data AugmentationMusic GenerationText-to-Music Generation