Fuse It or Lose It: Deep Fusion for Multimodal Simulation-Based Inference
We present multimodal neural posterior estimation (MultiNPE), a method to integrate heterogeneous data from different sources in simulation-based inference with neural networks. Inspired by advances in deep fusion, it allows researchers to analyze data from different domains and infer the parameters of complex mathematical models with increased accuracy. We consider three fusion approaches for MultiNPE (early, late, hybrid) and evaluate their performance in three challenging experiments. MultiNPE not only outperforms single-source baselines on a reference task, but also achieves superior inference on scientific models from cognitive neuroscience and cardiology. We systematically investigate the impact of partially missing data on the different fusion strategies. Across our experiments, late and hybrid fusion techniques emerge as the methods of choice for practical applications of multimodal simulation-based inference.
Code (1)
Similar Papers 제목 키워드 기반
SceneDiffuser: Efficient and Controllable Driving Simulation Initialization and Rollout
Realistic and interactive scene simulation is a key prerequisite for autonomous vehicle (AV) development. In this work, we present SceneDiffuser, a scene-level diffusion prior designed for traffic simulation. It offers a…
DenoisingLarge Language ModelScene GenerationFUSE: FK-Steered Multi-Modal Flow Matching for Efficient Simulation-Based Posterior Estimation
Simulation-Based Inference (SBI) is critical for scientific discovery, with generative models offering a promising path toward efficient inference. However, existing methods struggle with effective multimodal modeling. T…
Using Diffusion Ensembles to Estimate Uncertainty for End-to-End Autonomous Driving
End-to-end planning systems for autonomous driving are improving rapidly, especially in closed-loop simulation environments like CARLA. Many such driving systems either do not consider uncertainty as part of the plan its…
Autonomous DrivingCARLA longest6Trajectory PlanningDynamic Multimodal Fusion
Deep multimodal learning has achieved great progress in recent years. However, current fusion approaches are static in nature, i.e., they process and fuse multimodal inputs with identical computation, without accounting …
Computational EfficiencySemantic SegmentationSentiment AnalysisMultimodal Fusion of EMG and Vision for Human Grasp Intent Inference in Prosthetic Hand Control
Objective: For transradial amputees, robotic prosthetic hands promise to regain the capability to perform daily living activities. Current control methods based on physiological signals such as electromyography (EMG) are…
Electroencephalogram (EEG)Electromyography (EMG)