paper-with-me

Papers

Pathway-Structured Privileged Distillation for Deployable Computational Pathology

2026-06-01 · Yongxin Guo, Hao Lu, Onur Koyun, Muhammet Demir, Metin Gurcan arxiv

Integrating transcriptomics and histopathology can improve cancer risk modelling, yet practical use is constrained by the limited availability of RNA profiling in routine settings. Here we introduce Mixture of Pathway Experts (MoPE), a knowledge-distillation framework that reframes multimodal learning as privileged distillation for histology-only inference. MoPE is motivated by the partial observability between RNA profiles and whole-slide images: histology can capture morphology-linked consequences of certain molecular programmes, but cannot be expected to reconstruct the full transcriptomic state. MoPE encodes RNA-derived pathways and transfers the molecular supervision to pathway-indexed pathology experts through memory-usage alignment. Across diverse public benchmarks and two independent breast cancer cohorts, MoPE consistently improved WSI-only inference performance relative to baseline methods. Pathway-usage analyses and human-audited visual inspection provide bounded inspection of model behaviour and candidate morphology-linked readouts. These results support pathway-structured privileged distillation as a promising route to using molecular information during training while preserving RNA-free inference.

📄 PDF Abstract BibTeX arXiv:2606.02877

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Latent On-Policy Self-Distillation

2026-08-13 · Guibin Zhang, Jiayang Lyu, Ran Sun, Xinlei Yu 외 arxiv

Enabling agents to learn from experience and internalize it into their policy has become a central problem in self-evolving AI. On-policy self-distillation (OPSD) offers an effective pathway by using a privileged self-te…

Code Generation

Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning

2026-06-17 · Pengyu Li, Zhitao Gao, Lingling Zhang, Muye Huang 외 arxiv

Unified multimodal models (UMMs) interleave generated ''visual thoughts'' (VTs) with text reasoning to improve spatial tasks. This incurs roughly an order-of-magnitude inference cost from multi-step diffusion. We find th…

Multimodal Reasoning

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization

2026-06-05 · Shizhe Xiang, Ke An, Wenlong Yu, Yue Liu 외 arxiv

Recent post-training methods, particularly Reinforcement Learning with Verifiable Rewards (RLVR), have significantly enhanced the reasoning ability of Large Vision-Language Models (LVLMs). However, the sparse nature of v…

Reinforcement LearningMultimodal Reasoning

TAR: Teacher-Aligned Representations via Contrastive Learning for Quadrupedal Locomotion

2025-03-26 · Amr Mousa, Neil Karavis, Michele Caprio, Wei Pan 외

Quadrupedal locomotion via Reinforcement Learning (RL) is commonly addressed using the teacher-student paradigm, where a privileged teacher guides a proprioceptive student policy. However, key challenges such as represen…

Contrastive LearningReinforcement Learning (RL)TAR

SCDP: Learning Humanoid Locomotion from Partial Observations via Mixed-Observation Distillation

2026-03-10 · Milo Carroll, Tianhu Peng, Lingfan Bao, Chengxu Zhou 외 arxiv

Distilling humanoid locomotion control from offline datasets into deployable policies remains a challenge, as existing methods rely on privileged full-body states that require complex and often unreliable state estimatio…