paper-with-me

홈 › Papers

Curriculum-enhanced GroupDRO: Challenging the Norm of Avoiding Curriculum Learning in Subpopulation Shift Setups

2024-11-22 · Antonio Barbalau

In subpopulation shift scenarios, a Curriculum Learning (CL) approach would only serve to imprint the model weights, early on, with the easily learnable spurious correlations featured. To the best of our knowledge, none of the current state-of-the-art subpopulation shift approaches employ any kind of curriculum. To overcome this, we design a CL approach aimed at initializing the model weights in an unbiased vantage point in the hypothesis space which sabotages easy convergence towards biased hypotheses during the final optimization based on the entirety of the available data. We hereby propose a Curriculum-enhanced Group Distributionally Robust Optimization (CeGDRO) approach, which prioritizes the hardest bias-confirming samples and the easiest bias-conflicting samples, leveraging GroupDRO to balance the initial discrepancy in terms of difficulty. We benchmark our proposed method against the most popular subpopulation shift datasets, showing an increase over the state-of-the-art results across all scenarios, up to 6.2% on Waterbirds.

📄 PDF Abstract BibTeX arXiv:2411.15272

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Automatic Curriculum Learning with Gradient Reward Signals

2023-12-21 · Ryan Campbell, Junsang Yoon

This paper investigates the impact of using gradient norm reward signals in the context of Automatic Curriculum Learning (ACL) for deep reinforcement learning (DRL). We introduce a framework where the teacher model, util…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning

EGDCL: An Adaptive Curriculum Learning Framework for Unbiased Glaucoma Diagnosis

2020-08-01 · ECCV 2020 8 · Rongchang Zhao, Xuanlin Chen, Zailiang Chen, Shuo Li

Today's computer-aided diagnosis (CAD) model is still far from the clinical practice of glaucoma detection, mainly due to the training bias originating from 1) the normal-abnormal class imbalance and 2) the rare but sign…

Specificity

Reasoner for Real-World Event Detection: Scaling Reinforcement Learning via Adaptive Perplexity-Aware Sampling Strategy

2025-07-02 · Xiaoyun Zhang, Jingqing Ruan, Xing Ma, Yawen Zhu 외 arxiv

Detecting abnormal events in real-world customer service dialogues is highly challenging due to the complexity of business data and the dynamic nature of customer interactions. Moreover, models must demonstrate strong ou…

Reinforcement LearningAnomaly Detection

Rethinking Normalization Placement for LLMs: Post-Norm under Curriculum Depth Growing

2026-08-13 · Sheng Ren, Yadong Wang, Naiqiang Tan, Jiangang Kong 외 arxiv

Pre-norm is the standard normalization placement in modern Transformers because it facilitates joint optimization of full-depth models. We ask whether this preference persists when depth is introduced through a curriculu…

$L_{2,1}$-Norm Regularized Quaternion Matrix Completion Using Sparse Representation and Quaternion QR Decomposition

2023-09-07 · Juan Han, Kit Ian Kou, Jifei Miao, LiZhi Liu 외

Color image completion is a challenging problem in computer vision, but recent research has shown that quaternion representations of color images perform well in many areas. These representations consider the entire colo…

Matrix Completion