paper-with-me

홈 › Papers

Multitask Multimodal Fusion with Tabular Foundation Models for Peak and Durability Prediction of Pertussis Booster Response

2026-05-13 · Divya Sitani arxiv

Pertussis booster vaccination produces immune responses that vary widely across individuals in both peak magnitude and long-term durability. These two phases are governed by partly distinct biological compartments:peak reflects acute B-cell activation and antibody secretion, while durability reflects the establishment of long-term humoral memory. Yet most computational models target only one, missing the full boost-and-wane trajectory. Jointly predicting both is non-trivial because the two endpoints are biologically dissociated rather than redundant; samples are small, modalities are heterogeneous with structured missingness, and the two tasks rely on different measurement windows. We propose a multi-task contrastive multimodal fusion architecture combining frozen TabPFN-v2 per-modality encoders, a dual-label supervised contrastive loss that treats two subjects as a positive pair if they agree on the Task 1 label or the Task 2 label, modality dropout calibrated to empirical missingness, and missingness-masked attention fusion. Applied to a curated subset of the CMI-PB pertussis booster dataset (n = 158 subjects, four modalities, 44.9% with at least one modality missing; Spearman r = -0.58 between peak and durability, n = 96), the model achieves test AUROC 0.797 (95% CI [0.621, 0.948]) for peak response and 0.755 (95% CI [0.519, 0.945]) for durability, with both significant under joint label permutation (N = 1000; p = 0.002 and p = 0.045). Across logistic regression, XGBoost, and MLP baselines on raw features and on TabPFN embeddings, the proposed model is the only one whose 95% CIs lie above chance on both tasks simultaneously. Per-modality contribution analyses recover task-specific modality contributions consistent with the underlying immunology: peak prediction is carried by cytokine signatures, while durability is carried by baseline antibody features.

📄 PDF Abstract BibTeX arXiv:2605.12852

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multimodal Metadata Assignment for Cultural Heritage Artifacts

2024-06-01 · Luis Rei, Dunja Mladenić, Mareike Dorozynski, Franz Rottensteiner 외

We develop a multimodal classifier for the cultural heritage domain using a late fusion approach and introduce a novel dataset. The three modalities are Image, Text, and Tabular data. We based the image classifier on a R…

MultiTab: A Scalable Foundation for Multitask Learning on Tabular Data

2025-11-13 · Dimitrios Sinodinos, Jack Yi Wei, Narges Armanfard arxiv

Tabular data is the most abundant data type in the world, powering systems in finance, healthcare, e-commerce, and beyond. As tabular datasets grow and span multiple related targets, there is an increasing need to exploi…

Recommendation Systems

Multitask-Informed Prior for In-Context Learning on Tabular Data: Application to Steel Property Prediction

2026-03-24 · Dimitrios Sinodinos, Bahareh Nikpour, Jack Yi Wei, Sushant Sinha 외 arxiv

Accurate prediction of mechanical properties of steel during hot rolling processes, such as Thin Slab Direct Rolling (TSDR), remains challenging due to complex interactions among chemical compositions, processing paramet…

Computational Efficiency

Many Heads but One Brain: Fusion Brain -- a Competition and a Single Multimodal Multitask Architecture

2021-11-22 · Daria Bakshandaeva, Denis Dimitrov, Vladimir Arkhipkin, Alex Shonenkov 외

Supporting the current trend in the AI community, we present the AI Journey 2021 Challenge called Fusion Brain, the first competition which is targeted to make the universal architecture which could process different mod…

Handwritten Text Recognitionobject-detectionObject DetectionQuestion Answering+4

OmniVec2 - A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning

2024-01-01 · CVPR 2024 1 · Siddharth Srivastava, Gaurav Sharma

We present a novel multimodal multitask network and associated training algorithm. The method is capable of ingesting data from approximately 12 different modalities namely image video audio text depth point cloud ti…

3D Point Cloud ClassificationAction ClassificationAction RecognitionAudio Classification+6