paper-with-me

Papers

AdParaphrase: Paraphrase Dataset for Analyzing Linguistic Features toward Generating Attractive Ad Texts

2025-02-07 · Soichiro Murakami, Peinan Zhang, Hidetaka Kamigaito, Hiroya Takamura, Manabu Okumura

Effective linguistic choices that attract potential customers play crucial roles in advertising success. This study aims to explore the linguistic features of ad texts that influence human preferences. Although the creation of attractive ad texts is an active area of research, progress in understanding the specific linguistic features that affect attractiveness is hindered by several obstacles. First, human preferences are complex and influenced by multiple factors, including their content, such as brand names, and their linguistic styles, making analysis challenging. Second, publicly available ad text datasets that include human preferences are lacking, such as ad performance metrics and human feedback, which reflect people's interests. To address these problems, we present AdParaphrase, a paraphrase dataset that contains human preferences for pairs of ad texts that are semantically equivalent but differ in terms of wording and style. This dataset allows for preference analysis that focuses on the differences in linguistic features. Our analysis revealed that ad texts preferred by human judges have higher fluency, longer length, more nouns, and use of bracket symbols. Furthermore, we demonstrate that an ad text-generation model that considers these findings significantly improves the attractiveness of a given text. The dataset is publicly available at: https://github.com/CyberAgentAILab/AdParaphrase.

📄 PDF Abstract BibTeX arXiv:2502.04674

Code (1)

cyberagentailab/adparaphrase 공식 구현

Tasks

Text Generation

Similar Papers 제목 키워드 기반

BnPC: A Corpus for Paraphrase Detection in Bangla

2021-12-17 · ACL ARR December 2022 12 · Anonymous

In this paper, we present the first benchmark dataset for paraphrase detection in Bangla language. Despite being the sixth most spoken language in the world, paraphrase identification in the Bangla language is barely ex…

Paraphrase IdentificationSentence

Paraphrase Types Elicit Prompt Engineering Capabilities

2024-06-28 · Jan Philip Wahle, Terry Ruas, Yang Xu, Bela Gipp

Much of the success of modern language models depends on finding a suitable prompt to instruct the model. Until now, it has been largely unknown how variations in the linguistic expression of prompts affect these models.…

DiversityPrompt Engineering

Towards Human Understanding of Paraphrase Types in ChatGPT

2024-07-02 · Dominik Meier, Jan Philip Wahle, Terry Ruas, Bela Gipp

Paraphrases represent a human's intuitive ability to understand expressions presented in various different ways. Current paraphrase evaluations of language models primarily use binary approaches, offering limited interpr…

Sentence

A Multi-cascaded Model with Data Augmentation for Enhanced Paraphrase Detection in Short Texts

2019-12-27 · Muhammad Haroon Shakeel, Asim Karim, Imdadullah Khan

Paraphrase detection is an important task in text analytics with numerous applications such as plagiarism detection, duplicate question identification, and enhanced customer support helpdesks. Deep models have been propo…

Data Augmentation

Paraphrase Types for Generation and Detection

2023-10-23 · Jan Philip Wahle, Bela Gipp, Terry Ruas

Current approaches in paraphrase generation and detection heavily rely on a single general similarity score, ignoring the intricate linguistic properties of language. This paper introduces two new tasks to address this s…

Binary ClassificationParaphrase Generation