paper-with-me

Papers

Contrastive Learning from Exploratory Actions: Leveraging Natural Interactions for Preference Elicitation

2025-01-02 · Nathaniel Dennler, Stefanos Nikolaidis, Maja Matarić

People have a variety of preferences for how robots behave. To understand and reason about these preferences, robots aim to learn a reward function that describes how aligned robot behaviors are with a user's preferences. Good representations of a robot's behavior can significantly reduce the time and effort required for a user to teach the robot their preferences. Specifying these representations -- what "features" of the robot's behavior matter to users -- remains a difficult problem; Features learned from raw data lack semantic meaning and features learned from user data require users to engage in tedious labeling processes. Our key insight is that users tasked with customizing a robot are intrinsically motivated to produce labels through exploratory search; they explore behaviors that they find interesting and ignore behaviors that are irrelevant. To harness this novel data source of exploratory actions, we propose contrastive learning from exploratory actions (CLEA) to learn trajectory features that are aligned with features that users care about. We learned CLEA features from exploratory actions users performed in an open-ended signal design activity (N=25) with a Kuri robot, and evaluated CLEA features through a second user study with a different set of users (N=42). CLEA features outperformed self-supervised features when eliciting user preferences over four metrics: completeness, simplicity, minimality, and explainability.

📄 PDF Abstract BibTeX arXiv:2501.01367

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learning

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Social Interactions Clustering MOOC Students: An Exploratory Study

2020-08-10 · Lei Shi, Alexandra Cristea, Ahmad Alamri, Armando M. Toda 외

An exploratory study on social interactions of MOOC students in FutureLearn was conducted, to answer "how can we cluster students based on their social interactions?" Comments were categorized based on how students inter…

BIG-bench Machine LearningClustering

Atom-Motif Contrastive Transformer for Molecular Property Prediction

2023-10-11 · Wentao Yu, Shuo Chen, Chen Gong, Gang Niu 외

Recently, Graph Transformer (GT) models have been widely used in the task of Molecular Property Prediction (MPP) due to their high reliability in characterizing the latent relationship among graph nodes (i.e., the atoms …

Molecular Property PredictionPredictionProperty Prediction

ProCeedRL: Process Critic with Exploratory Demonstration Reinforcement Learning for LLM Agentic Reasoning

2026-04-02 · Jingyue Gao, Yanjiang Guo, Xiaoshuai Chen, Jianyu Chen arxiv

Reinforcement Learning (RL) significantly enhances the reasoning abilities of large language models (LLMs), yet applying it to multi-turn agentic tasks remains challenging due to the long-horizon nature of interactions a…

Reinforcement Learning

What to align in multimodal contrastive learning?

2024-09-11 · Benoit Dufumier, Javiera Castillo-Navarro, Devis Tuia, Jean-Philippe Thiran

Humans perceive the world through multisensory integration, blending the information of different modalities to adapt their behavior. Contrastive learning offers an appealing solution for multimodal self-supervised learn…

Contrastive LearningSelf-Supervised Learning

Cooperative Design Optimization through Natural Language Interaction

2025-08-22 · Ryogo Niwa, Shigeo Yoshida, Yuki Koyama, Yoshitaka Ushiku arxiv

Designing successful interactions requires identifying optimal design parameters. To do so, designers often conduct iterative user testing and exploratory trial-and-error. This involves balancing multiple objectives in a…