paper-with-me

Papers

GoferBot: A Visual Guided Human-Robot Collaborative Assembly System

2023-04-18 · Zheyu Zhuang, Yizhak Ben-Shabat, Jiahao Zhang, Stephen Gould, Robert Mahony

The current transformation towards smart manufacturing has led to a growing demand for human-robot collaboration (HRC) in the manufacturing process. Perceiving and understanding the human co-worker's behaviour introduces challenges for collaborative robots to efficiently and effectively perform tasks in unstructured and dynamic environments. Integrating recent data-driven machine vision capabilities into HRC systems is a logical next step in addressing these challenges. However, in these cases, off-the-shelf components struggle due to generalisation limitations. Real-world evaluation is required in order to fully appreciate the maturity and robustness of these approaches. Furthermore, understanding the pure-vision aspects is a crucial first step before combining multiple modalities in order to understand the limitations. In this paper, we propose GoferBot, a novel vision-based semantic HRC system for a real-world assembly task. It is composed of a visual servoing module that reaches and grasps assembly parts in an unstructured multi-instance and dynamic environment, an action recognition module that performs human action prediction for implicit communication, and a visual handover module that uses the perceptual understanding of human behaviour to produce an intuitive and efficient collaborative assembly experience. GoferBot is a novel assembly system that seamlessly integrates all sub-modules by utilising implicit semantic information purely from visual perception.

📄 PDF Abstract BibTeX arXiv:2304.08840

Code (0)

등록된 구현이 없습니다.

Tasks

Action Recognition

Similar Papers 제목 키워드 기반

Creativity and Visual Communication from Machine to Musician: Sharing a Score through a Robotic Camera

2024-09-09 · Ross Greer, Laura Fleig, Shlomo Dubnov

This paper explores the integration of visual communication and musical interaction by implementing a robotic camera within a "Guided Harmony" musical game. We aim to examine co-creative behaviors between human musicians…

Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely

2026-05-29 · Chalamalasetti Kranti, Sherzod Hakimov, David Schlangen arxiv

Robots operating in diverse environments rely on visual input to interpret objects and spatial layouts. In human-collaborative tasks, they are expected to communicate this understanding through language. Vision-language …

Instruction FollowingQuestion AnsweringSpatial Reasoning

Semantic Segmentation of Underwater Imagery: Dataset and Benchmark

2020-04-02 · Md Jahidul Islam, Chelsey Edge, Yuyang Xiao, Peigen Luo 외

In this paper, we present the first large-scale dataset for semantic Segmentation of Underwater IMagery (SUIM). It contains over 1500 images with pixel annotations for eight object categories: fish (vertebrates), reefs (…

Computational EfficiencyDecoderSaliency PredictionScene Understanding+2

ExpressMM: Expressive Mobile Manipulation Behaviors in Human-Robot Interactions

2026-04-07 · Souren Pashangpour, Haitong Wang, Matthew Lisondra, Goldie Nejat arxiv

Mobile manipulators are increasingly deployed in human-centered environments to perform tasks. While completing such tasks, they should also be able to communicate their intent to the people around them using expressive …

Analysis of Mutual and Referential Human and Robot Gazes in a Collaborative Word Association Game

2026-07-13 · Jens V. Rüppel, Tim Schreiter, Andrey Rudenko, Achim J. Lilienthal arxiv

Robot gaze is a major component of human-robot dialogue coordination. Most studies of gaze in human-robot dialogue focus on face-to-face social conversations, but little is known about gaze in demanding task-focused inte…