Data-Driven Knowledge Transfer in Batch $Q^*$ Learning
In data-driven decision-making in marketing, healthcare, and education, it is desirable to utilize a large amount of data from existing ventures to navigate high-dimensional feature spaces and address data scarcity in new ventures. We explore knowledge transfer in dynamic decision-making by concentrating on batch stationary environments and formally defining task discrepancies through the lens of Markov decision processes (MDPs). We propose a framework of Transferred Fitted $Q$-Iteration algorithm with general function approximation, enabling the direct estimation of the optimal action-state function $Q^*$ using both target and source data. We establish the relationship between statistical performance and MDP task discrepancy under sieve approximation, shedding light on the impact of source and target sample sizes and task discrepancy on the effectiveness of knowledge transfer. We show that the final learning error of the $Q^*$ function is significantly improved from the single task rate both theoretically and empirically.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingMarketingNavigateQ-LearningTransfer LearningSimilar Papers 제목 키워드 기반
Momentum Memory for Knowledge Distillation in Computational Pathology
Multimodal learning that integrates genomics and histopathology has shown strong potential in cancer diagnosis, yet its clinical translation is hindered by the limited availability of paired histology-genomics data. Know…
Knowledge DistillationDesign of Dynamic Experiments: A Data-Driven Methodology for the Optimization of Time-Varying Processes
ABSTRACT: A new methodology for the design of experiments is presented that provides a way to optimize the operation of a variety of batch and semibatch or fed-batch processes without the use of a knowledge-driven or fu…
Transferability analysis of data-driven additive manufacturing knowledge: a case study between powder bed fusion and directed energy deposition
Data-driven research in Additive Manufacturing (AM) has gained significant success in recent years. This has led to a plethora of scientific literature to emerge. The knowledge in these works consists of AM and Artificia…
Transfer LearningResidential Demand Response Applications Using Batch Reinforcement Learning
Driven by recent advances in batch Reinforcement Learning (RL), this paper contributes to the application of batch RL to demand response. In contrast to conventional model-based approaches, batch RL techniques do not req…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)pFedDSH: Enabling Knowledge Transfer in Personalized Federated Learning through Data-free Sub-Hypernetwork
Federated Learning (FL) enables collaborative model training across distributed clients without sharing raw data, offering a significant privacy benefit. However, most existing Personalized Federated Learning (pFL) metho…
Personalized Federated LearningContinual Learning