paper-with-me

Papers

Optimizing Query Evaluations using Reinforcement Learning for Web Search

2018-04-12 · Corby Rosset, Damien Jose, Gargi Ghosh, Bhaskar Mitra, Saurabh Tiwary

In web search, typically a candidate generation step selects a small set of documents---from collections containing as many as billions of web pages---that are subsequently ranked and pruned before being presented to the user. In Bing, the candidate generation involves scanning the index using statically designed match plans that prescribe sequences of different match criteria and stopping conditions. In this work, we pose match planning as a reinforcement learning task and observe up to 20% reduction in index blocks accessed, with small or no degradation in the quality of the candidate sets.

📄 PDF Abstract BibTeX arXiv:1804.04410

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

QP-OneModel: A Unified Generative LLM for Multi-Task Query Understanding in Xiaohongshu Search

2026-02-10 · Jianzhao Huang, Xiaorui Huang, Fei Zhao, Yunpeng Liu 외 arxiv

Query Processing (QP) bridges user intent and content supply in large-scale Social Network Service (SNS) search engines. Traditional QP systems rely on pipelines of isolated discriminative models (e.g., BERT), suffering …

Reinforcement Learning

Annotation-Free Reinforcement Learning Query Rewriting via Verifiable Search Reward

2025-07-31 · Sungguk Cha, DongWook Kim, Taeseung Hahn, Mintae Kim 외 arxiv

Optimizing queries for Retrieval-Augmented Generation (RAG) systems poses a significant challenge, particularly across diverse modal indices. We introduce RL-QR, a novel annotation-free reinforcement learning framework f…

Reinforcement Learning

Fast Bayesian Optimization of Function Networks with Partial Evaluations

2025-06-13 · Poompol Buathong, Peter I. Frazier

Bayesian optimization of function networks (BOFN) is a framework for optimizing expensive-to-evaluate objective functions structured as networks, where some nodes' outputs serve as inputs for others. Many real-world appl…

Bayesian OptimizationDrug Discovery

DEEPRUBRIC: Evidence-Tree Rubric Supervision for Efficient Reinforcement Learning of Deep Research Agents

2026-06-15 · Minghang Zhu, Chuyang Wei, Junhao Xu, Yilin Cheng 외 arxiv

Deep research agents synthesize long-form reports by searching and reasoning over retrieved evidence. Reinforcement learning with rubric-based rewards improves these agents by optimizing them against checkable criteria t…

Reinforcement Learning

Semantic Equivalence of e-Commerce Queries

2023-08-07 · Aritra Mandal, Daniel Tunkelang, Zhe Wu

Search query variation poses a challenge in e-commerce search, as equivalent search intents can be expressed through different queries with surface-level differences. This paper introduces a framework to recognize and le…

SentenceSentence Similarity