paper-with-me

Papers

Paraphrase Acquisition from Image Captions

2023-01-26 · Marcel Gohsen, Matthias Hagen, Martin Potthast, Benno Stein

We propose to use image captions from the Web as a previously underutilized resource for paraphrases (i.e., texts with the same "message") and to create and analyze a corresponding dataset. When an image is reused on the Web, an original caption is often assigned. We hypothesize that different captions for the same image naturally form a set of mutual paraphrases. To demonstrate the suitability of this idea, we analyze captions in the English Wikipedia, where editors frequently relabel the same image for different articles. The paper introduces the underlying mining technology, the resulting Wikipedia-IPC dataset, and compares known paraphrase corpora with respect to their syntactic and semantic paraphrase similarity to our new resource. In this context, we introduce characteristic maps along the two similarity dimensions to identify the style of paraphrases coming from different sources. An annotation study demonstrates the high reliability of the algorithmically determined characteristic maps.

📄 PDF Abstract BibTeX arXiv:2301.11030

Code (1)

webis-de/eacl-23 공식 구현

Tasks

ArticlesImage Captioning

Similar Papers 제목 키워드 기반

Generating Diverse and Descriptive Image Captions Using Visual Paraphrases

2019-10-01 · ICCV 2019 10 · Lixin Liu, Jiajun Tang, Xiaojun Wan, Zongming Guo

Recently there has been significant progress in image captioning with the help of deep learning. However, captions generated by current state-of-the-art models are still far from satisfactory, despite high scores in term…

DescriptiveDiversityImage Captioning

Language-Guided Invariance Probing of Vision-Language Models

2025-11-17 · Jae Joong Lee arxiv

Recent vision-language models (VLMs) such as CLIP, OpenCLIP, EVA02-CLIP and SigLIP achieve strong zero-shot performance, but it is unclear how reliably they respond to controlled linguistic perturbations. We introduce La…

Image-text matching

Fine-tuning CLIP Text Encoders with Two-step Paraphrasing

2024-02-23 · Hyunjae Kim, Seunghyun Yoon, Trung Bui, Handong Zhao 외

Contrastive language-image pre-training (CLIP) models have demonstrated considerable success across various vision-language tasks, such as text-to-image retrieval, where the model is required to effectively process natur…

Image CaptioningImage RetrievalParaphrase GenerationRetrieval+1

MIPA: Mutual Information Based Paraphrase Acquisition via Bilingual Pivoting

2017-11-01 · IJCNLP 2017 11 · Tomoyuki Kajiwara, Mamoru Komachi, Daichi Mochihashi

We present a pointwise mutual information (PMI)-based approach to formalize paraphrasability and propose a variant of PMI, called MIPA, for the paraphrase acquisition. Our paraphrase acquisition method first acquires lex…

Learning Word EmbeddingsSemantic Textual SimilarityWord AlignmentWord Embeddings

Chinese Whispers: Cooperative Paraphrase Acquisition

2012-05-01 · LREC 2012 5 · Matteo Negri, Yashar Mehdad, Aless Marchetti, ro 외

We present a framework for the acquisition of sentential paraphrases based on crowdsourcing. The proposed method maximizes the lexical divergence between an original sentence s and its valid paraphrases by running a sequ…

Machine TranslationNatural Language InferenceQuestion AnsweringSentence+2