paper-with-me

Papers

Harnessing the Continuous Structure: Utilizing the First-order Approach in Online Contract Design

2024-03-11 · Shiliang Zuo

This work studies the online contract design problem. The principal's goal is to learn the optimal contract that maximizes her utility through repeated interactions, without prior knowledge of the agent's type (i.e., the agent's cost and production functions). We leverage the structure provided by continuous action spaces, which allows the application of first-order conditions (FOC) to characterize the agent's behavior. In some cases, we utilize conditions from the first-order approach (FOA) in economics, but in certain settings, we are able to apply FOC without additional assumptions, leading to simpler and more principled algorithms. We illustrate this approach in three problem settings. Firstly, we study the problem of learning the optimal contract when there can be many outcomes. In contrast to prior works that design highly specialized algorithms, we show that the problem can be directly reduced to Lipschitz bandits. Secondly, we study the problem of learning linear contracts. While the contracting problem involves hidden action (moral hazard) and the pricing problem involves hidden value (adverse selection), the two problems share a similar optimization structure, which enables direct reduction between the problem of learning linear contracts and dynamic pricing. Thirdly, we study the problem of learning contracts with many outcomes when agents are identical and provide an algorithm with polynomial sample complexity.

📄 PDF Abstract BibTeX arXiv:2403.07143

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Harnessing omnipresent oscillator networks as computational resource

2025-02-07 · Thomas Geert de Jong, Hirofumi Notsu, Kohei Nakajima

Nature is pervaded with oscillatory dynamics. In networks of coupled oscillators patterns can arise when the system synchronizes to an external input. Hence, these networks provide processing and memory of input. We pres…

AgentLTL: A Trace-Verification Framework for Measuring, Enforcing, and Training Procedural Compliance in Tool-Using LLM Agents

2026-07-01 · Laïla Elkoussy, Julien Perez arxiv

Tool-using LLM agents are usually evaluated by final-answer correctness or LLM judges. Neither captures how an answer was produced. In safety-critical settings, the procedure itself is part of correctness. In this paper,…

Bayesian Optimization over Hybrid Spaces

2021-06-08 · Aryan Deshwal, Syrine Belakaria, Janardhan Rao Doppa

We consider the problem of optimizing hybrid structures (mixture of discrete and continuous input variables) via expensive black-box function evaluations. This problem arises in many real-world applications. For example,…

Bayesian Optimization

Policy Optimization with Second-Order Advantage Information

2018-05-09 · Jiajin Li, Baoxiang Wang

Policy optimization on high-dimensional continuous control tasks exhibits its difficulty caused by the large variance of the policy gradient estimators. We present the action subspace dependent gradient (ASDG) estimator …

continuous-controlContinuous ControlMuJoCo

Harnessing Deep Neural Networks with Logic Rules

2016-03-21 · ACL 2016 8 · Zhiting Hu, Xuezhe Ma, Zhengzhong Liu, Eduard Hovy 외

Combining deep neural networks with structured logic rules is desirable to harness flexibility and reduce uninterpretability of the neural models. We propose a general framework capable of enhancing various types of neur…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Sentiment Analysis