paper-with-me

홈 › Papers

Pose-aware Multi-level Feature Network for Human Object Interaction Detection

2019-09-18 · ICCV 2019 10 · Bo Wan, Desen Zhou, Yongfei Liu, Rongjie Li, Xuming He

Reasoning human object interactions is a core problem in human-centric scene understanding and detecting such relations poses a unique challenge to vision systems due to large variations in human-object configurations, multiple co-occurring relation instances and subtle visual difference between relation categories. To address those challenges, we propose a multi-level relation detection strategy that utilizes human pose cues to capture global spatial configurations of relations and as an attention mechanism to dynamically zoom into relevant regions at human part level. Specifically, we develop a multi-branch deep network to learn a pose-augmented relation representation at three semantic levels, incorporating interaction context, object features and detailed semantic part cues. As a result, our approach is capable of generating robust predictions on fine-grained human object interactions with interpretable outputs. Extensive experimental evaluations on public benchmarks show that our model outperforms prior methods by a considerable margin, demonstrating its efficacy in handling complex scenes.

📄 PDF Abstract BibTeX arXiv:1909.08453

Code (1)

bobwan1995/PMFNet 공식 구현 pytorch

Tasks

Human-Object Interaction DetectionObjectRelationScene Understanding

Similar Papers 제목 키워드 기반

SHERF: Generalizable Human NeRF from a Single Image

2023-03-22 · ICCV 2023 1 · Shoukang Hu, Fangzhou Hong, Liang Pan, Haiyi Mei 외

Existing Human NeRF methods for reconstructing 3D humans typically rely on multiple 2D images from multi-view cameras or monocular videos captured from fixed camera views. However, in real-world scenarios, human images a…

3D Human ReconstructionNeRF

Towards Complex Backgrounds: A Unified Difference-Aware Decoder for Binary Segmentation

2022-10-27 · Jiepan Li, wei he, Hongyan zhang

Binary segmentation is used to distinguish objects of interest from background, and is an active area of convolutional encoder-decoder network research. The current decoders are designed for specific objects based on the…

Decoder

VoteHMR: Occlusion-Aware Voting Network for Robust 3D Human Mesh Recovery from Partial Point Clouds

2021-10-17 · Guanze Liu, Yu Rong, Lu Sheng

3D human mesh recovery from point clouds is essential for various tasks, including AR/VR and human behavior understanding. Previous works in this field either require high-quality 3D human scans or sequential point cloud…

Human Mesh Recovery

Progressive Limb-Aware Virtual Try-On

2025-03-16 · Xiaoyu Han, Shengping Zhang, Qinglin Liu, Zonglin Li 외

Existing image-based virtual try-on methods directly transfer specific clothing to a human image without utilizing clothing attributes to refine the transferred clothing geometry and textures, which causes incomplete and…

AttributeHuman ParsingVirtual Try-on

MP-Mat: A 3D-and-Instance-Aware Human Matting and Editing Framework with Multiplane Representation

2025-04-20 · Siyi Jiao, Wenzheng Zeng, Yerong Li, Huayu Zhang 외

Human instance matting aims to estimate an alpha matte for each human instance in an image, which is challenging as it easily fails in complex cases requiring disentangling mingled pixels belonging to multiple instances …

Image Matting