paper-with-me

홈 › Papers

The Value of AI-Generated Metadata for UGC Platforms: Evidence from a Large-scale Field Experiment

2024-12-24 · Xinyi Zhang, Chenshuo Sun, Renyu Zhang, Khim-Yong Goh

AI-generated content (AIGC), such as advertisement copy, product descriptions, and social media posts, is becoming ubiquitous in business practices. However, the value of AI-generated metadata, such as titles, remains unclear on user-generated content (UGC) platforms. To address this gap, we conducted a large-scale field experiment on a leading short-video platform in Asia to provide about 1 million users access to AI-generated titles for their uploaded videos. Our findings show that the provision of AI-generated titles significantly boosted content consumption, increasing valid watches by 1.6% and watch duration by 0.9%. When producers adopted these titles, these increases jumped to 7.1% and 4.1%, respectively. This viewership-boost effect was largely attributed to the use of this generative AI (GAI) tool increasing the likelihood of videos having a title by 41.4%. The effect was more pronounced for groups more affected by metadata sparsity. Mechanism analysis revealed that AI-generated metadata improved user-video matching accuracy in the platform's recommender system. Interestingly, for a video for which the producer would have posted a title anyway, adopting the AI-generated title decreased its viewership on average, implying that AI-generated titles may be of lower quality than human-generated ones. However, when producers chose to co-create with GAI and significantly revised the AI-generated titles, the videos outperformed their counterparts with either fully AI-generated or human-generated titles, showcasing the benefits of human-AI co-creation. This study highlights the value of AI-generated metadata and human-AI metadata co-creation in enhancing user-content matching and content consumption for UGC platforms.

📄 PDF Abstract BibTeX arXiv:2412.18337

Code (0)

등록된 구현이 없습니다.

Tasks

Recommendation Systems

Similar Papers 제목 키워드 기반

Show me the material evidence: Initial experiments on evaluating hypotheses from user-generated multimedia data

2016-11-11 · Bernardo Gonçalves

Subjective questions such as `does neymar dive', or `is clinton lying', or `is trump a fascist', are popular queries to web search engines, as can be seen by autocompletion suggestions on Google, Yahoo and Bing. In the e…

Leveraging User-Generated Metadata of Online Videos for Cover Song Identification

2024-12-16 · Simon Hachmeier, Robert Jäschke

YouTube is a rich source of cover songs. Since the platform itself is organized in terms of videos rather than songs, the retrieval of covers is not trivial. The field of cover song identification addresses this problem …

Cover song identificationEntity ResolutionRetrieval

Croissant Baker: Metadata Generation for Discoverable, Governable, and Reusable ML Datasets

2026-05-14 · Rafi Al Attrach, Rajna Fani, Sebastian Lobentanzer, Joan Giner-Miguelez 외 arxiv

Croissant has emerged as the metadata standard for machine learning datasets, providing a structured, JSON-LD-based format that makes dataset discovery, automated ingestion, and reproducible analysis machine-checkable ac…

Raison d’être of the benchmark dataset: A Survey of Current Practices of Benchmark Dataset Sharing Platforms

2022-05-01 · nlppower (ACL) 2022 5 · Jaihyun Park, Sullam Jeoung

This paper critically examines the current practices of benchmark dataset sharing in NLP and suggests a better way to inform reusers of the benchmark dataset. As the dataset sharing platform plays a key role not only in …

PIPER: Content-Based Table Search via profiling and LLM-Generated Pseudoqueries

2026-05-18 · Riccardo Terrenzi, Matteo Falconi, Serkan Ayvaz, Pierluigi Plebani arxiv

The rapid growth of tabular datasets in data lakes, data spaces, and open data portals makes effective dataset search essential for reuse and analysis. Existing search systems rely mainly on metadata, which is often inco…

Question Answering