paper-with-me

홈 › Papers

Exploring the Learning Difficulty of Data Theory and Measure

2022-05-16 · Weiyao Zhu, Ou wu, Fengguang Su, Yingjun Deng

As learning difficulty is crucial for machine learning (e.g., difficulty-based weighting learning strategies), previous literature has proposed a number of learning difficulty measures. However, no comprehensive investigation for learning difficulty is available to date, resulting in that nearly all existing measures are heuristically defined without a rigorous theoretical foundation. In addition, there is no formal definition of easy and hard samples even though they are crucial in many studies. This study attempts to conduct a pilot theoretical study for learning difficulty of samples. First, a theoretical definition of learning difficulty is proposed on the basis of the bias-variance trade-off theory on generalization error. Theoretical definitions of easy and hard samples are established on the basis of the proposed definition. A practical measure of learning difficulty is given as well inspired by the formal definition. Second, the properties for learning difficulty-based weighting strategies are explored. Subsequently, several classical weighting methods in machine learning can be well explained on account of explored properties. Third, the proposed measure is evaluated to verify its reasonability and superiority in terms of several main difficulty factors. The comparison in these experiments indicates that the proposed measure significantly outperforms the other measures throughout the experiments.

📄 PDF Abstract BibTeX arXiv:2205.07427

Code (1)

weiyao619/geld 공식 구현 pytorch

Similar Papers 제목 키워드 기반

From Kinetic Theory to AI: a Rediscovery of High-Dimensional Divergences and Their Properties

2025-07-15 · Gennaro Auricchio, Giovanni Brigati, Paolo Giudici, Giuseppe Toscani arxiv

Selecting an appropriate divergence measure is a critical aspect of machine learning, as it directly impacts model performance. Among the most widely used, we find the Kullback-Leibler (KL) divergence, originally introdu…

Exploring Properties of Icosoku by Constraint Satisfaction Approach

2019-08-16 · Ke Liu, Sven Löffler, Petra Hofstedt

Icosoku is a challenging and interesting puzzle that exhibits highly symmetrical and combinatorial nature. In this paper, we pose the questions derived from the puzzle, but with more difficulty and generality. In additio…

RIDE: Difficulty Evolving Perturbation with Item Response Theory for Mathematical Reasoning

2025-11-06 · Xinyuan Li, Murong Xu, Wenbiao Tao, Hanlun Zhu 외 arxiv

Large language models (LLMs) achieve high performance on mathematical reasoning, but these results can be inflated by training data leakage or superficial pattern matching rather than genuine reasoning. To this end, an a…

Reinforcement LearningMathematical Reasoning

Talent or Luck? Evaluating Attribution Bias in Large Language Models

2025-05-28 · Chahat Raj, Mahika Banerjee, Aylin Caliskan, Antonios Anastasopoulos 외

When a student fails an exam, do we tend to blame their effort or the test's difficulty? Attribution, defined as how reasons are assigned to event outcomes, shapes perceptions, reinforces stereotypes, and influences deci…

Fairness

Surprisal Theory is Tautological (without Rational Grounding)

2026-07-23 · Ryan Cotterell arxiv

Surprisal theory holds that the human processing difficulty of a linguistic unit in context is an affine function of its surprisal under some language model. I argue this claim is a tautology without further constraint: …