paper-with-me

홈 › Papers

Integrating Multimodal Data for Joint Generative Modeling of Complex Dynamics

2022-12-15 · Manuel Brenner, Florian Hess, Georgia Koppe, Daniel Durstewitz

Many, if not most, systems of interest in science are naturally described as nonlinear dynamical systems. Empirically, we commonly access these systems through time series measurements. Often such time series may consist of discrete random variables rather than continuous measurements, or may be composed of measurements from multiple data modalities observed simultaneously. For instance, in neuroscience we may have behavioral labels in addition to spike counts and continuous physiological recordings. While by now there is a burgeoning literature on deep learning for dynamical systems reconstruction (DSR), multimodal data integration has hardly been considered in this context. Here we provide such an efficient and flexible algorithmic framework that rests on a multimodal variational autoencoder for generating a sparse teacher signal that guides training of a reconstruction model, exploiting recent advances in DSR training techniques. It enables to combine various sources of information for optimal reconstruction, even allows for reconstruction from symbolic data (class labels) alone, and connects different types of observations within a common latent dynamics space. In contrast to previous multimodal data integration techniques for scientific applications, our framework is fully \textit{generative}, producing, after training, trajectories with the same geometrical and temporal structure as those of the ground truth system.

📄 PDF Abstract BibTeX arXiv:2212.07892

Code (2)

durstewitzlab/mtf 공식 구현 pytorch
DurstewitzLab/CNS-2023

Tasks

Data IntegrationTime SeriesTime Series Analysis

Similar Papers 제목 키워드 기반

What if Agents Could Imagine? Reinforcing Open-Vocabulary HOI Comprehension through Generation

2026-02-12 · Zhenlong Yuan, Yue Wang, Dapeng Zhang, Kejin Cui 외 arxiv

Multimodal Large Language Models have shown promising capabilities in bridging visual and textual reasoning, yet their reasoning capabilities in Open-Vocabulary Human-Object Interaction (OV-HOI) are limited by cross-moda…

Reinforcement LearningImage Cropping

A Multimodal Learning Framework for Comprehensive 3D Mineral Prospectivity Modeling with Jointly Learned Structure-Fluid Relationships

2023-09-06 · Yang Zheng, Hao Deng, Ruisheng Wang, Jingjie Wu

This study presents a novel multimodal fusion model for three-dimensional mineral prospectivity mapping (3D MPM), effectively integrating structural and fluid information through a deep network architecture. Leveraging C…

Data IntegrationDecision Making

Text-to-Image Generation Via Energy-Based CLIP

2024-08-30 · Roy Ganz, Michael Elad

Joint Energy Models (JEMs), while drawing significant research attention, have not been successfully scaled to real-world, high-resolution datasets. We present EB-CLIP, a novel approach extending JEMs to the multimodal v…

Image GenerationText to Image GenerationText-to-Image Generation

CoVAE: correlated multimodal generative modeling

2026-03-02 · Federico Caretti, Guido Sanguinetti arxiv

Multimodal Variational Autoencoders have emerged as a popular tool to extract effective representations from rich multimodal data. However, such models rely on fusion strategies in latent space that destroy the joint sta…

Score-Based Multimodal Autoencoder

2023-05-25 · Daniel Wesego, Pedram Rooshenas

Multimodal Variational Autoencoders (VAEs) represent a promising group of generative models that facilitate the construction of a tractable posterior within the latent space given multiple modalities. Previous studies ha…