paper-with-me

홈 › Papers

How to Learn when Data Gradually Reacts to Your Model

2021-12-13 · Zachary Izzo, James Zou, Lexing Ying

A recent line of work has focused on training machine learning (ML) models in the performative setting, i.e. when the data distribution reacts to the deployed model. The goal in this setting is to learn a model which both induces a favorable data distribution and performs well on the induced distribution, thereby minimizing the test loss. Previous work on finding an optimal model assumes that the data distribution immediately adapts to the deployed model. In practice, however, this may not be the case, as the population may take time to adapt to the model. In many applications, the data distribution depends on both the currently deployed ML model and on the "state" that the population was in before the model was deployed. In this work, we propose a new algorithm, Stateful Performative Gradient Descent (Stateful PerfGD), for minimizing the performative loss even in the presence of these effects. We provide theoretical guarantees for the convergence of Stateful PerfGD. Our experiments confirm that Stateful PerfGD substantially outperforms previous state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2112.07042

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

How to Learn when Data Reacts to Your Model: Performative Gradient Descent

2021-02-15 · Zachary Izzo, Lexing Ying, James Zou

Performative distribution shift captures the setting where the choice of which ML model is deployed changes the data distribution. For example, a bank which uses the number of open credit lines to determine a customer's …

You Only Use Reactive Attention Slice For Long Context Retrieval

2024-09-03 · Yun Joon Soh, Hanxian Huang, Yuandong Tian, Jishen Zhao

Supporting longer context for Large Language Models (LLM) is a promising direction to advance LLMs. As training a model for a longer context window is computationally expensive, many alternative solutions, such as Retrie…

RAGRetrievalRetrieval-augmented GenerationSentence

Trust the Model Where It Trusts Itself -- Model-Based Actor-Critic with Uncertainty-Aware Rollout Adaption

2024-05-29 · Bernd Frauenknecht, Artur Eisele, Devdutt Subhasish, Friedrich Solowjow 외

Dyna-style model-based reinforcement learning (MBRL) combines model-free agents with predictive transition models through model-based rollouts. This combination raises a critical question: 'When to trust your model?'; i.…

modelModel-based Reinforcement LearningMuJoCo

Realistic overground gait transitions are not sharp but involve gradually changing walk-run mixtures as per energy optimality

2025-01-01 · Nicholas S. Baker, Leroy Long, Manoj Srinivasan

Humans use two qualitatively different gaits for locomotion, namely, walking and running -- usually using walking at lower speeds and running at higher speeds. Researchers have examined when humans switch between walking…

A Wizard-of-Oz Study on A Non-Task-Oriented Dialog Systems That Reacts to User Engagement

2016-09-01 · WS 2016 9 · Zhou Yu, Leah Nicolich-Henkin, Alan W. black, Alex Rudnicky 외
Machine Translation