paper-with-me

홈 › Papers

Coupling User Preference with External Rewards to Enable Driver-centered and Resource-aware EV Charging Recommendation

2022-10-23 · Chengyin Li, Zheng Dong, Nathan Fisher, Dongxiao Zhu

Electric Vehicle (EV) charging recommendation that both accommodates user preference and adapts to the ever-changing external environment arises as a cost-effective strategy to alleviate the range anxiety of private EV drivers. Previous studies focus on centralized strategies to achieve optimized resource allocation, particularly useful for privacy-indifferent taxi fleets and fixed-route public transits. However, private EV driver seeks a more personalized and resource-aware charging recommendation that is tailor-made to accommodate the user preference (when and where to charge) yet sufficiently adaptive to the spatiotemporal mismatch between charging supply and demand. Here we propose a novel Regularized Actor-Critic (RAC) charging recommendation approach that would allow each EV driver to strike an optimal balance between the user preference (historical charging pattern) and the external reward (driving distance and wait time). Experimental results on two real-world datasets demonstrate the unique features and superior performance of our approach to the competing methods.

📄 PDF Abstract BibTeX arXiv:2210.12693

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Language Model Personalization via Reward Factorization

2025-03-08 · Idan Shenfeld, Felix Faltings, Pulkit Agrawal, Aldo Pacchiano

Modern large language models (LLMs) are optimized for human-aligned responses using Reinforcement Learning from Human Feedback (RLHF). However, existing RLHF approaches assume a universal preference model and fail to acc…

Language ModelingLanguage Modellingmodel

YFPO: Yoked Feature Preference Optimization with Neuron-Guided Rewards

2026-05-12 · Yifan Le arxiv

Preference optimization has become a widely used post-training paradigm for improving the reasoning abilities of large language models. Existing methods typically learn from preferred and dispreferred responses as extern…

Mathematical ReasoningLogical Reasoning

Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards

2026-03-01 · Seungwook Kim, Minsu Cho arxiv

Text-to-image generation powers content creation across design, media, and data augmentation. Post-training of text-to-image generative models is a promising path to improve human preference alignment, factuality, and ae…

Text-to-Image GenerationReinforcement LearningData Augmentation

Hierarchical Conversational Preference Elicitation with Bandit Feedback

2022-09-06 · Jinhang Zuo, Songwen Hu, Tong Yu, Shuai Li 외

The recent advances of conversational recommendations provide a promising way to efficiently elicit users' preferences via conversational interactions. To achieve this, the recommender system conducts conversations with …

Recommendation Systems

Self-Improving Diffusion Classifiers with Minority Preference Optimization

2026-07-04 · Hyunsoo Kim, Jungmyung Wi, Soobin Um, Donghyun Kim 외 arxiv

Prior studies have demonstrated that diffusion classifiers achieve robust zero-shot classification performance. However, their effectiveness is strongly tied to the pretraining data distribution: they perform well in maj…