paper-with-me

홈 › Papers

Advancing TDFN: Precise Fixation Point Generation Using Reconstruction Differences

2025-01-26 · Shuguang Wang, Yuanjing Wang

Wang and Wang (2025) proposed the Task-Driven Fixation Network (TDFN) based on the fixation mechanism, which leverages low-resolution information along with high-resolution details near fixation points to accomplish specific visual tasks. The model employs reinforcement learning to generate fixation points. However, training reinforcement learning models is challenging, particularly when aiming to generate pixel-level accurate fixation points on high-resolution images. This paper introduces an improved fixation point generation method by leveraging the difference between the reconstructed image and the input image to train the fixation point generator. This approach directs fixation points to areas with significant differences between the reconstructed and input images. Experimental results demonstrate that this method achieves highly accurate fixation points, significantly enhances the network's classification accuracy, and reduces the average number of required fixations to achieve a predefined accuracy level.

📄 PDF Abstract BibTeX arXiv:2501.15603

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

TDFNet: An Efficient Audio-Visual Speech Separation Model with Top-down Fusion

2024-01-25 · Samuel Pegg, Kai Li, Xiaolin Hu

Audio-visual speech separation has gained significant traction in recent years due to its potential applications in various fields such as speech recognition, diarization, scene analysis and assistive technologies. Desig…

speech-recognitionSpeech RecognitionSpeech Separation

Spatial statistics for gaze patterns in scene viewing: Effects of repeated viewing

2018-11-13

Scene viewing is used to study attentional selection in complex but still controlled environments. One of the main observations on eye movements during scene viewing is the inhomogeneous distribution of fixation location…

TDFNet: Tri-projection Deformable Fusion Network for Panoramic Salient Object Detection

2026-08-26 · Qiangqiang Zhou, Jiacong Yu, Jiawei Xu, Yong Chen 외 arxiv

Recent years have witnessed the growing potential of panoramic salient object detection in robotic vision, virtual reality, and related applications. However, projecting spherical scenes onto 2D planes inevitably introdu…

Salient Object Detection

Influence of initial fixation position in scene viewing

2016-07-13

During scene perception our eyes generate complex sequences of fixations. Predictors of fixation locations are bottom-up factors like luminance contrast, top-down factors like viewing instruction, and systematic biases l…

Position

An experimental and computational study of an Estonian single-person word naming

2025-09-03 · Kaidi Lõo, Arvi Tavast, Maria Heitmeier, Harald Baayen arxiv

This study investigates lexical processing in Estonian. A large-scale single-subject experiment is reported that combines the word naming task with eye-tracking. Five response variables (first fixation duration, total fi…