paper-with-me

Papers

Deep Occlusion Reasoning for Multi-Camera Multi-Target Detection

2017-04-19 · ICCV 2017 10 · Pierre Baqué, François Fleuret, Pascal Fua

People detection in single 2D images has improved greatly in recent years. However, comparatively little of this progress has percolated into multi-camera multi-people tracking algorithms, whose performance still degrades severely when scenes become very crowded. In this work, we introduce a new architecture that combines Convolutional Neural Nets and Conditional Random Fields to explicitly model those ambiguities. One of its key ingredients are high-order CRF terms that model potential occlusions and give our approach its robustness even when many people are present. Our model is trained end-to-end and we show that it outperforms several state-of-art algorithms on challenging scenes.

📄 PDF Abstract BibTeX arXiv:1704.05775

Code (2)

pierrebaque/DeepOcclusion
rickyHong/DeepOcclustion-repl

Tasks

Multiview Detection

Similar Papers 제목 키워드 기반

State-aware Re-identification Feature for Multi-target Multi-camera Tracking

2019-06-04 · Peng Li, Jiabin Zhang, Zheng Zhu, Yanwei Li 외

Multi-target Multi-camera Tracking (MTMCT) aims to extract the trajectories from videos captured by a set of cameras. Recently, the tracking performance of MTMCT is significantly enhanced with the employment of re-identi…

TrackVLA++: Unleashing Reasoning and Memory Capabilities in VLA Models for Embodied Visual Tracking

2025-10-08 · Jiahang Liu, Yunpeng Qi, Jiazhao Zhang, Minghan Li 외 arxiv

Embodied Visual Tracking (EVT) is a fundamental ability that underpins practical applications, such as companion robots, guidance robots and service assistants, where continuously following moving targets is essential. R…

Zero-shot GeneralizationSpatial ReasoningVisual Tracking

Multimodal Active Measurement for Human Mesh Recovery in Close Proximity

2023-10-12 · Takahiro Maeda, Keisuke Takeshita, Norimichi Ukita, Kazuhito Tanaka

For physical human-robot interactions (pHRI), a robot needs to estimate the accurate body pose of a target person. However, in these pHRI scenarios, the robot cannot fully observe the target person's body with equipped c…

Human Mesh RecoveryPose EstimationSensor Fusion

Hybrid Multi-camera Visual Servoing to Moving Target

2018-03-06 · Hanz Cuevas-Velasquez, Nanbo Li, Radim Tylecek, Marcelo Saval-Calvo 외

Visual servoing is a well-known task in robotics. However, there are still challenges when multiple visual sources are combined to accurately guide the robot or occlusions appear. In this paper we present a novel visual …

SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generation

2026-02-26 · Vaibhav Agrawal, Rishubh Parihar, Pradhaan Bhat, Ravi Kiran Sarvadevabhatla 외 arxiv

We identify occlusion reasoning as a fundamental yet overlooked aspect for 3D layout-conditioned generation. It is essential for synthesizing partially occluded objects with depth-consistent geometry and scale. While exi…

Text-to-Image Generation