paper-with-me

Papers Video Saliency Prediction

“Video Saliency Prediction” 태그가 달린 논문 29편 · 필터 해제

Text-Audio-Visual-conditioned Diffusion Model for Video Saliency Prediction

2025-04-19 · Li Yu, Xuanzhe Sun, Wei Zhou, Moncef Gabbouj

Video saliency prediction is crucial for downstream applications, such as video compression and human-computer interaction. With the flourishing of multimodal learning, researchers started to explore multimodal video sal…

DenoisingImage GenerationPredictionSaliency Prediction+2

DTFSal: Audio-Visual Dynamic Token Fusion for Video Saliency Prediction

2025-04-14 · Kiana Hooshanfar, Alireza Hosseini, Ahmad Kalhor, Babak Nadjar Araabi

Audio-visual saliency prediction aims to mimic human visual attention by identifying salient regions in videos through the integration of both visual and auditory information. Although visual-only approaches have signifi…

Computational EfficiencySaliency PredictionVideo Saliency Prediction

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues

2025-02-01 · Rohit Girmaji, Siddharth Jain, Bhav Beri, Sarthak Bansal 외

This paper introduces ViNet-S, a 36MB model based on the ViNet architecture with a U-Net design, featuring a lightweight decoder that significantly reduces model size and parameters without compromising performance. Addi…

Action ClassificationAction LocalizationDecoderSaliency Prediction+3

Relevance-guided Audio Visual Fusion for Video Saliency Prediction

2024-11-18 · Li Yu, Xuanzhe Sun, Pan Gao, Moncef Gabbouj

Audio data, often synchronized with video frames, plays a crucial role in guiding the audience's visual attention. Incorporating audio information into video saliency prediction tasks can enhance the prediction of human …

PredictionSaliency PredictionVideo Saliency Prediction

AIM 2024 Challenge on Video Saliency Prediction: Methods and Results

2024-09-23 · Andrey Moskalenko, Alexey Bryncev, Dmitry Vatolin, Radu Timofte 외

This paper reviews the Challenge on Video Saliency Prediction at AIM 2024. The goal of the participants was to develop a method for predicting accurate saliency maps for the provided set of video sequences. Saliency maps…

Saliency DetectionSaliency PredictionVideo CompressionVideo Saliency Detection+1

CaRDiff: Video Salient Object Ranking Chain of Thought Reasoning for Saliency Prediction with Diffusion

2024-08-21 · Yunlong Tang, Gen Zhan, Li Yang, Yiting Liao 외

Video saliency prediction aims to identify the regions in a video that attract human attention and gaze, driven by bottom-up features from the video and top-down processes like memory and cognition. Among these top-down …

Language ModellingLarge Language ModelMultimodal Large Language ModelPrediction+2

SalFoM: Dynamic Saliency Prediction with Video Foundation Models

2024-04-03 · Morteza Moradi, Mohammad Moradi, Francesco Rundo, Concetto Spampinato 외

Recent advancements in video saliency prediction (VSP) have shown promising performance compared to the human visual system, whose emulation is the primary goal of VSP. However, current state-of-the-art models employ spa…

DecoderPredictionSaliency PredictionVideo Saliency Prediction

Transformer-based Video Saliency Prediction with High Temporal Dimension Decoding

2024-01-15 · Morteza Moradi, Simone Palazzo, Concetto Spampinato

In recent years, finding an effective and efficient strategy for exploiting spatial and temporal information has been a hot research topic in video saliency prediction (VSP). With the emergence of spatio-temporal transfo…

DecoderSaliency PredictionVideo Saliency Prediction

UniST: Towards Unifying Saliency Transformer for Video Saliency Prediction and Detection

2023-09-15 · Junwen Xiong, Peng Zhang, Chuanyue Li, Wei Huang 외

Video saliency prediction and detection are thriving research domains that enable computers to simulate the distribution of visual attention akin to how humans perceiving dynamic scenes. While many approaches have crafte…

Decoderobject-detectionObject DetectionPrediction+4

Spherical Vision Transformer for 360-degree Video Saliency Prediction

2023-08-24 · Mert Cokelek, Nevrez Imamoglu, Cagri Ozcinar, Erkut Erdem 외

The growing interest in omnidirectional videos (ODVs) that capture the full field-of-view (FOV) has gained 360-degree saliency prediction importance in computer vision. However, predicting where humans look in 360-degree…

PredictionSaliency PredictionVideo Saliency PredictionVideo Understanding

CASP-Net: Rethinking Video Saliency Prediction from an Audio-VisualConsistency Perceptual Perspective

2023-03-11 · Junwen Xiong, Ganglai Wang, Peng Zhang, Wei Huang 외

Incorporating the audio stream enables Video Saliency Prediction (VSP) to imitate the selective attention mechanism of human brain. By focusing on the benefits of joint auditory and visual information, most VSP methods a…

DecoderSaliency PredictionVideo Saliency Prediction

TinyHD: Efficient Video Saliency Prediction with Heterogeneous Decoders using Hierarchical Maps Distillation

2023-01-11 · Feiyan Hu, Simone Palazzo, Federica Proietto Salanitri, Giovanni Bellitto 외

Video saliency prediction has recently attracted attention of the research community, as it is an upstream task for several practical applications. However, current solutions are particularly computationally demanding, e…

Knowledge DistillationPredictionSaliency PredictionVideo Saliency Prediction

CASP-Net: Rethinking Video Saliency Prediction From an Audio-Visual Consistency Perceptual Perspective

2023-01-01 · CVPR 2023 1 · Junwen Xiong, Ganglai Wang, Peng Zhang, Wei Huang 외

Incorporating the audio stream enables Video Saliency Prediction (VSP) to imitate the selective attention mechanism of human brain. By focusing on the benefits of joint auditory and visual information, most VSP metho…

DecoderSaliency PredictionVideo Saliency Prediction

GASP: Gated Attention For Saliency Prediction

2022-06-09 · International Joint Conference on Artificial Intelligence 2021 8 · Fares Abawi, Tom Weber, Stefan Wermter

Saliency prediction refers to the computational task of modeling overt attention. Social cues greatly influence our attention, consequently altering our eye movements and behavior. To emphasize the efficacy of such featu…

PredictionSaliency PredictionVideo Saliency DetectionVideo Saliency Prediction

Spatio-Temporal Self-Attention Network for Video Saliency Prediction

2021-08-24 · Ziqiang Wang, Zhi Liu, Gongyang Li, Yang Wang 외

3D convolutional neural networks have achieved promising results for video tasks in computer vision, including video saliency prediction that is explored in this paper. However, 3D convolution encodes visual representati…

PredictionSaliency PredictionVideo Saliency Prediction

Noise-Aware Video Saliency Prediction

2021-04-16 · Ekta Prashnani, Orazio Gallo, Joohwan Kim, Josef Spjut 외

We tackle the problem of predicting saliency maps for videos of dynamic scenes. We note that the accuracy of the maps reconstructed from the gaze data of a fixed number of observers varies with the frame, as it depends o…

PredictionSaliency PredictionVideo Saliency Prediction

ViNet: Pushing the limits of Visual Modality for Audio-Visual Saliency Prediction

2020-12-11 · Samyak Jain, Pradeep Yarlagadda, Shreyank Jyoti, Shyamgopal Karthik 외

We propose the ViNet architecture for audio-visual saliency prediction. ViNet is a fully convolutional encoder-decoder architecture. The encoder uses visual features from a network trained for action recognition, and the…

Action RecognitionDecoderPredictionSaliency Prediction+2

Hierarchical Domain-Adapted Feature Learning for Video Saliency Prediction

2020-10-02 · Giovanni Bellitto, Federica Proietto Salanitri, Simone Palazzo, Francesco Rundo 외

In this work, we propose a 3D fully convolutional architecture for video saliency prediction that employs hierarchical supervision on intermediate maps (referred to as conspicuity maps) generated using features extracted…

Domain AdaptationSaliency DetectionSaliency PredictionUnsupervised Domain Adaptation+2

Video Saliency Prediction Using Enhanced Spatiotemporal Alignment Network

2020-01-02 · Jin Chen, Huihui Song, Kaihua Zhang, Bo Liu 외

Due to a variety of motions across different frames, it is highly challenging to learn an effective spatiotemporal representation for accurate video saliency prediction (VSP). To address this issue, we develop an effecti…

PredictionSaliency PredictionVideo Saliency DetectionVideo Saliency Prediction

Simple vs complex temporal recurrences for video saliency prediction

2019-07-03 · Panagiotis Linardos, Eva Mohedano, Juan Jose Nieto, Noel E. O'Connor 외

This paper investigates modifying an existing neural network architecture for static saliency prediction using two types of recurrences that integrate information from the temporal domain. The first modification is the a…

PredictionSaliency PredictionVideo Saliency DetectionVideo Saliency Prediction
1–20 / 29 다음 →