paper-with-me

홈 › Papers

User-Feedback-Driven Adaptation for Vision-and-Language Navigation

2025-12-11 · Yongqiang Yu, Xuhui Li, Hazza Mahmood, Jinxing Zhou, Haodong Hong, Longtao Jiang, Zhiqiang Xu, Qi Wu, Xiaojun Chang arxiv

Real-world deployment of Vision-and-Language Navigation (VLN) agents is constrained by the scarcity of reliable supervision after offline training. While recent adaptation methods attempt to mitigate distribution shifts via environment-driven self-supervision (e.g., entropy minimization), these signals are often noisy and can cause the agent to amplify its own mistakes during long-horizon sequential decision-making. In this paper, we propose a paradigm shift that positions user feedback, specifically episode-level success confirmations and goal-level corrections, as a primary and general-purpose supervision signal for VLN. Unlike internal confidence scores, user feedback is intent-aligned and in-situ consistent, directly correcting the agent's decoupling from user instructions. To effectively leverage this supervision, we introduce a user-feedback-driven learning framework featuring a topology-aware trajectory construction pipeline. This mechanism lifts sparse, goal-level corrections into dense path-level supervision by generating feasible paths on the agent's incrementally built topological graph, enabling sample-efficient imitation learning without requiring step-by-step human demonstrations. Furthermore, we develop a persistent memory bank mechanism for warm-start initialization, supporting the reuse of previously acquired topology and cached representations across navigation sessions. Extensive experiments on the GSA-R2R benchmark demonstrate that our approach transforms sparse interaction into robust supervision, consistently outperforming environment-driven baselines while exhibiting strong adaptability across diverse instruction styles.

📄 PDF Abstract BibTeX arXiv:2512.10322

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

NExT-Search: Rebuilding User Feedback Ecosystem for Generative AI Search

2025-05-20 · Sunhao Dai, Wenjie Wang, Liang Pang, Jun Xu 외

Generative AI search is reshaping information retrieval by offering end-to-end answers to complex queries, reducing users' reliance on manually browsing and summarizing multiple web pages. However, while this paradigm en…

Answer GenerationInformation RetrievalRetrieval

Adaptive Dynamic Dehazing via Instruction-Driven and Task-Feedback Closed-Loop Optimization for Diverse Downstream Task Adaptation

2026-02-28 · Yafei Zhang, Shuaitian Song, Huafeng Li, Shujuan Wang 외 arxiv

In real-world vision systems,haze removal is required not only to enhance image visibility but also to meet the specific needs of diverse downstream tasks.To address this challenge,we propose a novel adaptive dynamic deh…

Lusifer: LLM-based User SImulated Feedback Environment for online Recommender systems

2024-05-22 · Danial Ebrat, Eli Paradalis, Luis Rueda

Reinforcement learning (RL) recommender systems often rely on static datasets that fail to capture the fluid, ever changing nature of user preferences in real-world scenarios. Meanwhile, generative AI techniques have eme…

Collaborative FilteringRecommendation Systemsreinforcement-learningReinforcement Learning+2

Past, Present, and Future of Bug Tracking in the Generative AI Era

2025-10-09 · Utku Boran Torun, Mehmet Taha Demircan, Mahmut Furkan Gön, Eray Tüzün arxiv

Traditional bug-tracking systems rely heavily on manual reporting, reproduction, classification, and resolution, involving multiple stakeholders such as end users, customer support, developers, and testers. This division…

Feedback Adaptation for Retrieval-Augmented Generation

2026-04-08 · Jihwan Bang, Seunghan Yang, Kyuhong Shim, Simyung Chang 외 arxiv

Retrieval-Augmented Generation (RAG) systems are typically evaluated under static assumptions, despite being frequently corrected through user or expert feedback in deployment. Existing evaluation protocols focus on over…