paper-with-me

홈 › Papers

Are Large Language Models Good Data Preprocessors?

2025-02-24 · Elyas Meguellati, Nardiena Pratama, Shazia Sadiq, Gianluca Demartini

High-quality textual training data is essential for the success of multimodal data processing tasks, yet outputs from image captioning models like BLIP and GIT often contain errors and anomalies that are difficult to rectify using rule-based methods. While recent work addressing this issue has predominantly focused on using GPT models for data preprocessing on relatively simple public datasets, there is a need to explore a broader range of Large Language Models (LLMs) and tackle more challenging and diverse datasets. In this study, we investigate the use of multiple LLMs, including LLaMA 3.1 70B, GPT-4 Turbo, and Sonnet 3.5 v2, to refine and clean the textual outputs of BLIP and GIT. We assess the impact of LLM-assisted data cleaning by comparing downstream-task (SemEval 2024 Subtask "Multilabel Persuasion Detection in Memes") models trained on cleaned versus non-cleaned data. While our experimental results show improvements when using LLM-cleaned captions, statistical tests reveal that most of these improvements are not significant. This suggests that while LLMs have the potential to enhance data cleaning and repairing, their effectiveness may be limited depending on the context they are applied to, the complexity of the task, and the level of noise in the text. Our findings highlight the need for further research into the capabilities and limitations of LLMs in data preprocessing pipelines, especially when dealing with challenging datasets, contributing empirical evidence to the ongoing discussion about integrating LLMs into data preprocessing pipelines.

📄 PDF Abstract BibTeX arXiv:2502.16790

Code (0)

등록된 구현이 없습니다.

Tasks

Image Captioning

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Attention 설명 없음
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Residual Connection 설명 없음

Similar Papers 제목 키워드 기반

Understanding Unconventional Preprocessors in Deep Convolutional Neural Networks for Face Identification

2019-03-27 · Chollette C. Olisah, Lyndon Smith

Deep networks have achieved huge successes in application domains like object and face recognition. The performance gain is attributed to different facets of the network architecture such as: depth of the convolutional l…

Data AugmentationFace IdentificationFace RecognitionQuantization+1

Preprocessors Matter! Realistic Decision-Based Attacks on Machine Learning Systems

2022-10-07 · Chawin Sitawarin, Florian Tramèr, Nicholas Carlini

Decision-based attacks construct adversarial examples against a machine learning (ML) model by making only hard-label queries. These attacks have mainly been applied directly to standalone neural networks. However, in pr…

Preprocessor Selection for Machine Learning Pipelines

2018-10-23 · Brandon Schoenfeld, Christophe Giraud-Carrier, Mason Poggemann, Jarom Christensen 외

Much of the work in metalearning has focused on classifier selection, combined more recently with hyperparameter optimization, with little concern for data preprocessing. Yet, it is generally well accepted that machine l…

BIG-bench Machine LearningHyperparameter Optimization

Computed tomography using meta-optics

2024-11-13 · Maksym Zhelyeznuyakov, Johannes E. Fröch, Shane Colburn, Steven L. Brunton 외

Computer vision tasks require processing large amounts of data to perform image classification, segmentation, and feature extraction. Optical preprocessors can potentially reduce the number of floating point operations r…

image-classificationImage ClassificationImage Reconstruction

Vector OFDM Transmission over Non-Gaussian Power Line Communication Channels

2018-06-26

Most of the recent power line communication (PLC) systems and standards, both narrow-band and broadband, are based on orthogonal frequency-division multiplexing (OFDM). This multiplexing scheme however suffers from the h…