Papers Occlusion Handling
“Occlusion Handling” 태그가 달린 논문 93편 · 필터 해제
Cross-Spectral Body Recognition with Side Information Embedding: Benchmarks on LLCM and Analyzing Range-Induced Occlusions on IJB-MDF
Vision Transformers (ViTs) have demonstrated impressive performance across a wide range of biometric tasks, including face and body recognition. In this work, we adapt a ViT model pretrained on visible (VIS) imagery to t…
Occlusion HandlingPerson Re-IdentificationGeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control
Recent advancements in world models have revolutionized dynamic environment simulation, allowing systems to foresee future states and assess potential actions. In autonomous driving, these capabilities help vehicles anti…
3D geometryAutonomous DrivingAutonomous NavigationOcclusion HandlingVTBench: Comprehensive Benchmark Suite Towards Real-World Virtual Try-on Models
While virtual try-on has achieved significant progress, evaluating these models towards real-world scenarios remains a challenge. A comprehensive benchmark is essential for three key reasons:(1) Current metrics inadequat…
Occlusion HandlingVirtual Try-onSelf-Supervised Monocular Visual Drone Model Identification through Improved Occlusion Handling
Ego-motion estimation is vital for drones when flying in GPS-denied environments. Vision-based methods struggle when flight speed increases and close-by objects lead to difficult visual conditions with considerable motio…
Motion EstimationOcclusion HandlingPose EstimationSelf-Supervised Learning+1MASSeg : 2nd Technical Report for 4th PVUW MOSE Track
Complex video object segmentation continues to face significant challenges in small object recognition, occlusion handling, and dynamic scene modeling. This report presents our solution, which ranked second in the MOSE t…
Data AugmentationObjectObject RecognitionOcclusion Handling+4Aligning Foundation Model Priors and Diffusion-Based Hand Interactions for Occlusion-Resistant Two-Hand Reconstruction
Two-hand reconstruction from monocular images faces persistent challenges due to complex and dynamic hand postures and occlusions, causing significant difficulty in achieving plausible interaction alignment. Existing app…
DenoisingOcclusion Handling8-Calves Image dataset
We introduce the 8-Calves dataset, a benchmark for evaluating object detection and identity classification in occlusion-rich, temporally consistent environments. The dataset comprises a 1-hour video (67,760 frames) of ei…
object-detectionObject DetectionOcclusion HandlingEMT: A Visual Multi-Task Benchmark Dataset for Autonomous Driving in the Arab Gulf Region
This paper introduces the Emirates Multi-Task (EMT) dataset - the first publicly available dataset for autonomous driving collected in the Arab Gulf region. The EMT dataset captures the unique road topology, high traffic…
Autonomous DrivingOcclusion HandlingTrajectory ForecastingOccludeNet: A Causal Journey into Mixed-View Actor-Centric Video Action Recognition under Occlusions
The lack of occlusion data in commonly used action recognition video datasets limits model robustness and impedes sustained performance improvements. We construct OccludeNet, a large-scale occluded video dataset that inc…
Action ClassificationAction RecognitionCausal Inferencecounterfactual+5NexusSplats: Efficient 3D Gaussian Splatting in the Wild
While 3D Gaussian Splatting (3DGS) has recently demonstrated remarkable rendering quality and efficiency in 3D scene reconstruction, it struggles with varying lighting conditions and incidental occlusions in real-world s…
3DGS3D Scene ReconstructionOcclusion HandlingIMUVIE: Pickup Timeline Action Localization via Motion Movies
Falls among seniors due to difficulties with tasks such as picking up objects pose significant health and safety risks, impacting quality of life and independence. Reliable, accessible assessment tools are critical for e…
Action ClassificationAction LocalizationOcclusion HandlingPCNet: a human pose compensation network based on incremental learning for sports actions estimation
Human pose estimation has a wide range of applications. Existing methods perform well in conventional domains, but there are certain defects when they are applied to sports activities. The first is lack of estimation of…
2D Human Pose EstimationIncremental LearningOcclusion HandlingPose EstimationA-MFST: Adaptive Multi-Flow Sparse Tracker for Real-Time Tissue Tracking Under Occlusion
Purpose: Tissue tracking is critical for downstream tasks in robot-assisted surgery. The Sparse Efficient Neural Depth and Deformation (SENDD) model has previously demonstrated accurate and real-time sparse point trackin…
Occlusion HandlingPoint TrackingDepGAN: Leveraging Depth Maps for Handling Occlusions and Transparency in Image Composition
Image composition is a complex task which requires a lot of information about the scene for an accurate and realistic composition, such as perspective, lighting, shadows, occlusions, and object interactions. Previous met…
Generative Adversarial NetworkOcclusion HandlingTransparent objectsVisual Multi-Object Tracking with Re-Identification and Occlusion Handling using Labeled Random Finite Sets
This paper proposes an online visual multi-object tracking (MOT) algorithm that resolves object appearance-reappearance and occlusion. Our solution is based on the labeled random finite set (LRFS) filtering approach, whi…
Multi-Object TrackingObjectObject TrackingOcclusion HandlingFeatureSORT: Essential Features for Effective Tracking
In this work, we introduce a novel tracker designed for online multiple object tracking with a focus on being simple, while being effective. we provide multiple feature modules each of which stands for a particular appea…
Multi-Object TrackingMultiple Object TrackingObject TrackingOcclusion HandlingCapeX: Category-Agnostic Pose Estimation from Textual Point Explanation
Conventional 2D pose estimation models are constrained by their design to specific object categories. This limits their applicability to predefined objects. To overcome these limitations, category-agnostic pose estimatio…
2D Pose EstimationAnimal Pose EstimationCategory-Agnostic Pose EstimationDecoder+6Track Initialization and Re-Identification for~3D Multi-View Multi-Object Tracking
We propose a 3D multi-object tracking (MOT) solution using only 2D detections from monocular cameras, which automatically initiates/terminates tracks as well as resolves track appearance-reappearance and occlusions. More…
3D Multi-Object TrackingMulti-Object TrackingObjectObject Tracking+1Occlusion Handling in 3D Human Pose Estimation with Perturbed Positional Encoding
Understanding human behavior fundamentally relies on accurate 3D human pose estimation. Graph Convolutional Networks (GCNs) have recently shown promising advancements, delivering state-of-the-art performance with rather …
3D Human Pose EstimationOcclusion HandlingPose EstimationCMU-Flownet: Exploring Point Cloud Scene Flow Estimation in Occluded Scenario
Occlusions hinder point cloud frame alignment in LiDAR data, a challenge inadequately addressed by scene flow models tested mainly on occlusion-free datasets. Attempts to integrate occlusion handling within networks ofte…
Occlusion EstimationOcclusion HandlingScene Flow Estimation