paper-with-me

홈 › Papers

Seeing Through Fog: Towards Fog-Invariant Action Recognition

2026-05-20 · Enqi Liu, Liyuan Pan, Zhi Gao, Lingzhi Li, Qing Li arxiv

Foggy conditions are commonly encountered in real-world applications; however, existing action recognition approaches typically assume favorable weather and high-quality video inputs. On foggy days, unpredictable visibility degradation and reduced contrast obstruct the extraction of semantic cues, posing significant challenges for current action recognition methods. In this paper, we mitigate the issues faced in action recognition under foggy conditions by employing two strategies. First, we present FogAct, the first benchmark dataset for foggy action recognition, consisting of paired clean and foggy videos captured with a stereo camera system. The dataset spans 10 scenes and 55 action categories, comprising nearly 10,000 video clips. Second, we propose FogNet, a two-stream CLIP model that discovers fog-invariant semantic information hidden behind the degraded videos. FogNet learns robust representations of foggy videos with guidance from clean videos, effectively capturing shared structural and motion cues between clean and foggy videos. Extensive experiments on FogAct and three other popular datasets demonstrate that our method achieves competitive performance compared with state-of-the-art (SOTA) approaches. Our FogAct and FogNet are given in our project page.

📄 PDF Abstract BibTeX arXiv:2605.20645

Code (0)

등록된 구현이 없습니다.

Tasks

Action Recognition

Similar Papers 제목 키워드 기반

Making the Invisible Visible: Action Recognition Through Walls and Occlusions

2019-09-20 · ICCV 2019 10 · Tianhong Li, Lijie Fan, Ming-Min Zhao, Yingcheng Liu 외

Understanding people's actions and interactions typically depends on seeing them. Automating the process of action recognition from visual data has been the topic of much research in the computer vision community. But wh…

3D Human Pose EstimationAction RecognitionRF-based Pose EstimationSkeleton Based Action Recognition

Vision: looking and seeing through our brain's information bottleneck

2025-03-24 · Li Zhaoping

Our brain recognizes only a tiny fraction of sensory input, due to an information processing bottleneck. This blinds us to most visual inputs. Since we are blind to this blindness, only a recent framework highlights this…

The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World

2023-08-03 · Weiyun Wang, Min Shi, Qingyun Li, Wenhai Wang 외

We present the All-Seeing (AS) project: a large-scale data and model for recognizing and understanding everything in the open world. Using a scalable data engine that incorporates human feedback and efficient models in t…

AllQuestion AnsweringRetrievalText Retrieval

Flip-Invariant Motion Representation

2017-10-01 · ICCV 2017 10 · Takumi Kobayashi

In action recognition, local motion descriptors contribute to effectively representing video sequences where target actions appear in localized spatio-temporal regions. For robust recognition, those fundamental descripto…

Action ClassificationAction RecognitionGeneral ClassificationTemporal Action Localization

Seeing by haptic glance: reinforcement learning-based 3D object Recognition

2021-02-15 · Kevin Riou, Suiyi Ling, Guillaume Gallot, Patrick Le Callet

Human is able to conduct 3D recognition by a limited number of haptic contacts between the target object and his/her fingers without seeing the object. This capability is defined as `haptic glance' in cognitive neuroscie…

3D Object RecognitionObjectObject Recognitionreinforcement-learning+2