paper-with-me

홈 › Papers

WPG-MoE: Weak-Prior-Guided Dense Mixture-of-Experts for User-Level Social Media Depression Detection

2026-07-05 · Xian Li, Yuanhe Tian, Yang Yang, Guoqing Wang, Yan Song arxiv

Online social media posts provide scalable signals for early depression screening, and recent studies mainly improve pre-classification evidence through risk-post selection, symptom grounding, and clinically informed feature construction. However, these screening-stage designs often leave final decisions to a single detector, overlooking how users heterogeneously express depressive risk after screening. A monolithic classifier must average across heterogeneous users, which may dilute localized evidence and cause misclassification, especially for non-self-disclosing users. To address this issue, we propose WPG-MoE, a weak-prior-guided dense mixture-of-experts framework built on a shared large language model (LLM) backbone. WPG-MoE derives user-level weak semantic priors to softly route users to experts matched to different evidence layouts. We formulate this process as learning using privileged information (LUPI): rich LLM-extracted structured evidence guides training-time routing, while inference retains only Patient Health Questionnaire-9 (PHQ-9) template screening and the deployable backbone. Experiments on Chinese and English datasets show that WPG-MoE outperforms strong baselines with interpretable routing behavior.

📄 PDF Abstract BibTeX arXiv:2607.04350

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Exposing and Defending the Achilles' Heel of Video Mixture-of-Experts

2026-02-01 · Songping Wang, Qinglong Liu, Yueming Lyu, Ning Li 외 arxiv

Mixture-of-Experts (MoE) has demonstrated strong performance in video understanding tasks, yet its adversarial robustness remains underexplored. Existing attack methods often treat MoE as a unified architecture, overlook…

Adversarial Robustness

AME-TS: Anchored Mixture-of-Experts for Time Series Forecasting

2026-05-24 · Rui Wang, Renhao Xue, Ray Razi, Huan Song 외 arxiv

Time series forecasting models are increasingly scaled through large Transformer backbones, yet most existing approaches process all series through a shared dense computation path despite substantial heterogeneity in tem…

Time Series Forecasting

Mixture of Experts Guided by Gaussian Splatters Matters: A new Approach to Weakly-Supervised Video Anomaly Detection

2025-08-08 · Giacomo D'Amicantonio, Snehashis Majhi, Quan Kong, Lorenzo Garattoni 외 arxiv

Video Anomaly Detection (VAD) is a challenging task due to the variability of anomalous events and the limited availability of labeled data. Under the Weakly-Supervised VAD (WSVAD) paradigm, only video-level labels are p…

Weakly-supervised Video Anomaly Detection

Pruning and Distilling Mixture-of-Experts into Dense Language Models

2026-05-27 · Junhyuck Kim, Jihun Yun, Haechan Kim, Gyeongman Kim 외 arxiv

Mixture-of-Experts (MoE) is now the dominant architecture for frontier language models, yet it requires all expert parameters to be loaded in memory, making it less preferable for memory-constrained deployment. Existing …

Knowledge Distillation

Domain-Expert-Guided Hybrid Mixture-of-Experts for Medical AI: Integrating Data-Driven Learning with Clinical Priors

2026-01-25 · Jinchen Gu, Nan Zhao, Lei Qiu, Lu Zhang arxiv

Mixture-of-Experts (MoE) models increase representational capacity with modest computational cost, but their effectiveness in specialized domains such as medicine is limited by small datasets. In contrast, clinical pract…