paper-with-me

홈 › Papers

Online Preselection with Context Information under the Plackett-Luce Model

2020-02-11 · Adil El Mesaoudi-Paul, Viktor Bengs, Eyke Hüllermeier

We consider an extension of the contextual multi-armed bandit problem, in which, instead of selecting a single alternative (arm), a learner is supposed to make a preselection in the form of a subset of alternatives. More specifically, in each iteration, the learner is presented a set of arms and a context, both described in terms of feature vectors. The task of the learner is to preselect $k$ of these arms, among which a final choice is made in a second step. In our setup, we assume that each arm has a latent (context-dependent) utility, and that feedback on a preselection is produced according to a Plackett-Luce model. We propose the CPPL algorithm, which is inspired by the well-known UCB algorithm, and evaluate this algorithm on synthetic and real data. In particular, we consider an online algorithm selection scenario, which served as a main motivation of our problem setting. Here, an instance (which defines the context) from a certain problem class (such as SAT) can be solved by different algorithms (the arms), but only $k$ of these algorithms can actually be run.

📄 PDF Abstract BibTeX arXiv:2002.04275

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Preselection Bandits

2019-07-13 · ICML 2020 1 · Viktor Bengs, Eyke Hüllermeier

In this paper, we introduce the Preselection Bandit problem, in which the learner preselects a subset of arms (choice alternatives) for a user, which then chooses the final arm from this subset. The learner is not aware …

CNN training with graph-based sample preselection: application to handwritten character recognition

2017-12-06 · Frédéric Rayar, Masanori Goto, Seiichi Uchida

In this paper, we present a study on sample preselection in large training data set for CNN-based classification. To do so, we structure the input data set in a network representation, namely the Relative Neighbourhood G…

General Classification

Online Rank Elicitation for Plackett-Luce: A Dueling Bandits Approach

2015-12-01 · NeurIPS 2015 12 · Balázs Szörényi, Róbert Busa-Fekete, Adil Paul, Eyke Hüllermeier

We study the problem of online rank elicitation, assuming that rankings of a set of alternatives obey the Plackett-Luce distribution. Following the setting of the dueling bandits problem, the learner is allowed to query …

Preselection via Classification: A Case Study on Evolutionary Multiobjective Optimization

2017-08-03 · Jinyuan Zhang, Aimin Zhou, Ke Tang, Guixu Zhang

In evolutionary algorithms, a preselection operator aims to select the promising offspring solutions from a candidate offspring set. It is usually based on the estimated or real objective values of the candidate offsprin…

ClassificationEvolutionary AlgorithmsGeneral ClassificationMultiobjective Optimization

PLR: Plackett-Luce for Reordering In-Context Learning Examples

2026-03-22 · Pawel Batorski, Paul Swoboda arxiv

In-context learning (ICL) adapts large language models by conditioning on a small set of ICL examples, avoiding costly parameter updates. Among other factors, performance is often highly sensitive to the ordering of the …

Mathematical Reasoning