paper-with-me

Papers

Multi-CALF: A Policy Combination Approach with Statistical Guarantees

2025-05-18 · Georgiy Malaniya, Anton Bolychev, Grigory Yaremenko, Anastasia Krasnaya, Pavel Osinenko

We introduce Multi-CALF, an algorithm that intelligently combines reinforcement learning policies based on their relative value improvements. Our approach integrates a standard RL policy with a theoretically-backed alternative policy, inheriting formal stability guarantees while often achieving better performance than either policy individually. We prove that our combined policy converges to a specified goal set with known probability and provide precise bounds on maximum deviation and convergence time. Empirical validation on control tasks demonstrates enhanced performance while maintaining stability guarantees.

📄 PDF Abstract BibTeX arXiv:2505.12350

Code (1)

aidagroup/multi-calf 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

A novel agent with formal goal-reaching guarantees: an experimental study with a mobile robot

2024-09-23 · Grigory Yaremenko, Dmitrii Dobriborsci, Roman Zashchitin, Ruben Contreras Maestre 외

Reinforcement Learning (RL) has been shown to be effective and convenient for a number of tasks in robotics. However, it requires the exploration of a sufficiently large number of state-action pairs, many of which may be…

Reinforcement Learning (RL)

HierarchicalForecast: A Reference Framework for Hierarchical Forecasting in Python

2022-07-07 · Kin G. Olivares, Azul Garza, David Luo, Cristian Challú 외

Large collections of time series data are commonly organized into structures with different levels of aggregation; examples include product and geographical groupings. It is often important to ensure that the forecasts a…

BIG-bench Machine LearningDecision MakingTime SeriesTime Series Analysis

Learning Social Robot Navigation By Sensing Human Legs

2026-07-30 · Alberto Vaglio, Andrea Garulli, Antonio Giannitrapani, Renato Quartullo 외 arxiv

Robots navigating among pedestrians typically sense their surroundings with a 2D LiDAR mounted close to the ground. At that height, the sensor mostly sees moving legs rather than whole people, yet most learning-based nav…

Reinforcement LearningRobot Navigation

FocalFormer3D: Focusing on Hard Instance for 3D Object Detection

2023-01-01 · ICCV 2023 1 · Yilun Chen, Zhiding Yu, Yukang Chen, Shiyi Lan 외

False negatives (FN) in 3D object detection, e.g., missing predictions of pedestrians, vehicles, or other obstacles, can lead to potentially dangerous situations in autonomous driving. While being fatal, this issue i…

3D Object DetectionAutonomous DrivingDecoderObject+2

Evaluating ROCKET and Catch22 features for calf behaviour classification from accelerometer data using Machine Learning models

2024-04-28 · Oshana Dissanayake, Sarah E. McPherson, Joseph Allyndree, Emer Kennedy 외

Monitoring calf behaviour continuously would be beneficial to identify routine practices (e.g., weaning, dehorning, etc.) that impact calf welfare in dairy farms. In that regard, accelerometer data collected from neck co…

Time SeriesTime Series Classification