paper-with-me

홈 › Papers

SkillSelect-Serve: QoS-Aware Budgeted Skill Service Recommendation for LLM Agents

2026-05-08 · Jingyuan Zheng, Dongjing Wang, Xin Zhang, Butian Huang, Haiping Zhang, Dongjin Yu, Shuguang Deng arxiv

Reusable agent skills are emerging as a service-oriented capability layer for Large Language Model (LLM) agents. Unlike plain retrieval items, a skill exposes functional capabilities, input-output assumptions, tool dependencies, context cost, and risk metadata. Selecting skills is particularly challenging for small LLM agents, which can load only a few capability units under restricted context, tool availability, and risk tolerance. Existing fixed Top-k methods rank skills by textual relevance and overlook requirement satisfaction, deliverability, and operational constraints. We present SkillSelect-Serve, a QoS-aware, budget-constrained Skill Service recommendation framework. Raw skills are profiled as structured Skill Services, the task is converted into a structured requirement object, and candidates discovered from a large-scale registry are ranked by a calibrated task-conditioned suitability estimator and packed by a constrained projection enforcing token-budget, aggregated-risk, and tool-availability constraints, using only deployment-observable features. On a registry of 35,353 skills with pooled multi-positive relevance judgments verified by two independent assessors, the unconstrained top-5 recommendation fits a realistic 4,000-token context for only 9.1% of tasks; the constrained projection restores 100% deliverability at a cost of only 1.14 points of hit rate, outperforming retrieve-and-rerank, budget truncation, and diversity-based selection under identical budgets. The same mechanism halves delivered risk exposure and eliminates the 44-81% tool-violation rates of tool-agnostic recommendation. At an identical three-service budget, hit rate improves from 0.8864 to 0.9091 over fixed Top-3 retrieval. The results support managing reusable agent skills as discoverable, comparable, and constraint-aware service units instead of plain retrievable documents.

📄 PDF Abstract BibTeX arXiv:2607.00011

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Graph-of-Skills: Dependency-Aware Structural Retrieval for Massive Agent Skills

2026-04-07 · Dawei Liu, Zongxia Li, Hongyang Du, Xiyang Wu 외 arxiv

Modern LLM agents increasingly rely on reusable skills, and as they interact with personal applications, web browsers, and other interfaces, skill libraries can scale to thousands of skills. Scaling to larger skill sets …

Semantic Retrieval

COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization

2026-09-10 · Pingchen Lu, Xiangyi Wang, Xiang Li, Jie Mao 외 hf

Large language model (LLM) agents can benefit from reusable skills distilled from prior task experience, yet existing skill optimization methods often rely on costly execution-based evaluation and substantial task data. …

Learning State-Dependent Policy Parametrizations for Dynamic Technician Routing with Rework

2024-09-03 · Jonas Stein, Florentin D Hildebrandt, Barrett W Thomas, Marlin W Ulmer

Home repair and installation services require technicians to visit customers and resolve tasks of different complexity. Technicians often have heterogeneous skills and working experiences. The geographical spread of cust…

Belief-Guided Inference Control for Large Language Model Services via Verifiable Observations

2026-04-30 · Wenhao Yuan, Chenchen Lin, Jian Chen, Jinfeng Xu 외 arxiv

In black-box large language model (LLM) services, response reliability is often only partially observable at decision time, while stronger inference pathways incur substantial computational cost, inducing a budgeted sequ…

Budgeted Training: Rethinking Deep Neural Network Training Under Resource Constraints

2019-05-12 · ICLR 2020 1 · Mengtian Li, Ersin Yumer, Deva Ramanan

In most practical settings and theoretical analyses, one assumes that a model can be trained until convergence. However, the growing complexity of machine learning datasets and models may violate such assumptions. Indeed…

General Classificationimage-classificationImage ClassificationInstance Segmentation+5