Multi-CALF: A Policy Combination Approach with Statistical Guarantees
We introduce Multi-CALF, an algorithm that intelligently combines reinforcement learning policies based on their relative value improvements. Our approach integrates a standard RL policy with a theoretically-backed alternative policy, inheriting formal stability guarantees while often achieving better performance than either policy individually. We prove that our combined policy converges to a specified goal set with known probability and provide precise bounds on maximum deviation and convergence time. Empirical validation on control tasks demonstrates enhanced performance while maintaining stability guarantees.
Code (1)
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
A novel agent with formal goal-reaching guarantees: an experimental study with a mobile robot
Reinforcement Learning (RL) has been shown to be effective and convenient for a number of tasks in robotics. However, it requires the exploration of a sufficiently large number of state-action pairs, many of which may be…
Reinforcement Learning (RL)HierarchicalForecast: A Reference Framework for Hierarchical Forecasting in Python
Large collections of time series data are commonly organized into structures with different levels of aggregation; examples include product and geographical groupings. It is often important to ensure that the forecasts a…
BIG-bench Machine LearningDecision MakingTime SeriesTime Series AnalysisLearning Social Robot Navigation By Sensing Human Legs
Robots navigating among pedestrians typically sense their surroundings with a 2D LiDAR mounted close to the ground. At that height, the sensor mostly sees moving legs rather than whole people, yet most learning-based nav…
Reinforcement LearningRobot NavigationFocalFormer3D: Focusing on Hard Instance for 3D Object Detection
False negatives (FN) in 3D object detection, e.g., missing predictions of pedestrians, vehicles, or other obstacles, can lead to potentially dangerous situations in autonomous driving. While being fatal, this issue i…
3D Object DetectionAutonomous DrivingDecoderObject+2Evaluating ROCKET and Catch22 features for calf behaviour classification from accelerometer data using Machine Learning models
Monitoring calf behaviour continuously would be beneficial to identify routine practices (e.g., weaning, dehorning, etc.) that impact calf welfare in dairy farms. In that regard, accelerometer data collected from neck co…
Time SeriesTime Series Classification