paper-with-me

홈 › Papers

Predicting Online Video Advertising Effects with Multimodal Deep Learning

2020-12-22 · Jun Ikeda, Hiroyuki Seshime, Xueting Wang, Toshihiko Yamasaki

With expansion of the video advertising market, research to predict the effects of video advertising is getting more attention. Although effect prediction of image advertising has been explored a lot, prediction for video advertising is still challenging with seldom research. In this research, we propose a method for predicting the click through rate (CTR) of video advertisements and analyzing the factors that determine the CTR. In this paper, we demonstrate an optimized framework for accurately predicting the effects by taking advantage of the multimodal nature of online video advertisements including video, text, and metadata features. In particular, the two types of metadata, i.e., categorical and continuous, are properly separated and normalized. To avoid overfitting, which is crucial in our task because the training data are not very rich, additional regularization layers are inserted. Experimental results show that our approach can achieve a correlation coefficient as high as 0.695, which is a significant improvement from the baseline (0.487).

📄 PDF Abstract BibTeX arXiv:2012.11851

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningMultimodal Deep Learning

Similar Papers 제목 키워드 기반

AgenticGen: Reward-Guided Agentic Video Generation for Advertising

2026-08-31 · Xingyuan Bu, Chengru Song, Hao Zhou, Tao Zhou 외 hf

Advertising video generation is not only a video synthesis task, but also a product-conditioned reasoning problem whose success is measured by online business metrics. Recent video foundation models can generate realisti…

Video Generation

ContextIQ: A Multimodal Expert-Based Video Retrieval System for Contextual Advertising

2024-10-29 · Ashutosh Chaubey, Anoubhav Agarwaal, Sartaki Sinha Roy, Aayush Agrawal 외

Contextual advertising serves ads that are aligned to the content that the user is viewing. The rapid growth of video content on social platforms and streaming services, along with privacy concerns, has increased the nee…

RetrievalText to Video RetrievalVideo Retrieval

Alignment Helps Make the Most of Multimodal Data

2024-05-14 · Christian Arnold, Andreas Küpfer

When studying political communication, combining the information from text, audio, and video signals promises to reflect the richness of human communication more comprehensively than confining it to individual modalities…

On the Effects of Video Grounding on Language Models

2022-10-01 · MMMPIE (COLING) 2022 10 · Ehsan Doostmohammadi, Marco Kuhlmann

Transformer-based models trained on text and vision modalities try to improve the performance on multimodal downstream tasks or tackle the problem Transformer-based models trained on text and vision modalities try to imp…

Image CaptioningQuestion AnsweringVideo GroundingVisual Question Answering+1

Online Causal Inference for Advertising in Real-Time Bidding Auctions

2019-08-22 · Caio Waisman, Harikesh S. Nair, Carlos Carrion

Real-time bidding (RTB) systems, which utilize auctions to allocate user impressions to competing advertisers, continue to enjoy success in digital advertising. Assessing the effectiveness of such advertising remains a c…

Causal InferenceExperimental DesignThompson Sampling