paper-with-me

Papers

MAPLE: Modality-Aware Post-training and Learning Ecosystem

2026-02-12 · Nikhil Verma, Minjung Kim, JooYoung Yoo, Kyung-Min Jin, Manasa Bharadwaj, Kevin Ferreira, Ko Keun Kim, Youngjoon Kim arxiv

Multimodal language models now integrate text, audio, and video for unified reasoning. Yet existing RL post-training pipelines treat all input signals as equally relevant, ignoring which modalities each task actually requires. This modality-blind training inflates policy-gradient variance, slows convergence, and degrades robustness to real-world distribution shifts where signals may be missing, added, or reweighted. We introduce MAPLE, a complete modality-aware post-training and learning ecosystem comprising: (1) MAPLE-bench, the first benchmark explicitly annotating minimal signal combinations required per task; (2) MAPO, a modality-aware policy optimization framework that stratifies batches by modality requirement to reduce gradient variance from heterogeneous group advantages; (3) Adaptive weighting and curriculum scheduling that balances and prioritizes harder signal combinations. Systematic analysis across loss aggregation, clipping, sampling, and curriculum design establishes MAPO's optimal training strategy. Adaptive weighting and curriculum focused learning further boost performance across signal combinations. MAPLE narrows uni/multi-modal accuracy gaps by 30.24%, converges 3.18x faster, and maintains stability across all modality combinations under realistic reduced signal access. MAPLE constitutes a complete recipe for deployment-ready multimodal RL post-training.

📄 PDF Abstract BibTeX arXiv:2602.11596

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MAPLE: Multi-Path Adaptive Propagation with Level-Aware Embeddings for Hierarchical Multi-Label Image Classification

2026-03-31 · Boshko Koloski, Marjan Stoimchev, Jurica Levatić, Dragi Kocev 외 arxiv

Hierarchical multi-label classification (HMLC) is essential for modeling structured label dependencies in remote sensing. Yet existing approaches struggle in multi-path settings, where images may activate multiple taxono…

Hierarchical Multi-label ClassificationMulti-Label Image Classification

MAPLE: A Mobile Assistant with Persistent Finite State Machines for Recovery Reasoning

2025-05-29 · Linqiang Guo, Wei Liu, Yi Wen Heng, Tse-Hsun 외

Mobile GUI agents aim to autonomously complete user-instructed tasks across mobile apps. Recent advances in Multimodal Large Language Models (MLLMs) enable these agents to interpret UI screens, identify actionable elemen…

MAPLE: Metadata Conditioned LLM Pretraining for Locale-Aware Question Answering

2026-01-21 · Anjishnu Mukherjee, Ziwei Zhu, Antonios Anastasopoulos arxiv

Large language models can memorize competing locale-specific facts yet fail to select among them when the locale changes, defaulting instead to a single globally dominant answer. We formalize this as localized knowledge …

Online Bayesian Moment Matching based SAT Solver Heuristics

2020-01-01 · ICML 2020 1 · Haonan Duan, Saeed Nejati, George Trimponias, Pascal Poupart 외

In this paper, we present a Bayesian Moment Matching (BMM) based method aimed at solving the initialization problem in Boolean SAT solvers. The initialization problem can be stated as follows: given a SAT formula φ, com…

MAPLE-X: Latency Prediction with Explicit Microprocessor Prior Knowledge

2022-05-25 · Saad Abbasi, Alexander Wong, Mohammad Javad Shafiee

Deep neural network (DNN) latency characterization is a time-consuming process and adds significant cost to Neural Architecture Search (NAS) processes when searching for efficient convolutional neural networks for embedd…

Neural Architecture SearchPrediction