paper-with-me

Papers

Grounding of the Functional Object-Oriented Network in Industrial Tasks

2022-04-05 · Rafik Ayari, Matteo Pantano, David Paulius

In this preliminary work, we propose to design an activity recognition system that is suitable for Industrie 4.0 (I4.0) applications, especially focusing on Learning from Demonstration (LfD) in collaborative robot tasks. More precisely, we focus on the issue of data exchange between an activity recognition system and a collaborative robotic system. We propose an activity recognition system with linked data using functional object-oriented network (FOON) to facilitate industrial use cases. Initially, we drafted a FOON for our use case. Afterwards, an action is estimated by using object and hand recognition systems coupled with a recurrent neural network, which refers to FOON objects and states. Finally, the detected action is shared via a context broker using an existing linked data model, thus enabling the robotic system to interpret the action and execute it afterwards. Our initial results show that FOON can be used for an industrial use case and that we can use existing linked data models in LfD applications.

📄 PDF Abstract BibTeX arXiv:2204.02274

Code (0)

등록된 구현이 없습니다.

Tasks

Activity Recognition

Similar Papers 제목 키워드 기반

Affordance2Action: Task-Conditioned Scene-level Affordance Grounding for Real-Time Manipulation

2026-06-02 · Litao Liu, Yifan Han, Pengfei Yi, Wenbo Yu 외 arxiv

Task-conditioned manipulation requires grounding instructions to task-relevant functional parts rather than object categories. This setting is scene-dependent and often one-to-many in cluttered scenes: the same object ma…

ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos

2025-12-03 · Qi'ao Xu, Tianwen Qian, Yuqian Fu, Kailing Li 외 arxiv

A core capability towards general embodied intelligence lies in localizing task-relevant objects from an egocentric perspective, formulated as Spatio-Temporal Video Grounding (STVG). Despite recent progress, existing STV…

Spatio-Temporal Video Grounding

Task-oriented Sequential Grounding in 3D Scenes

2024-08-07 · Zhuofan Zhang, Ziyu Zhu, Pengxiang Li, Tengyu Liu 외

Grounding natural language in physical 3D environments is essential for the advancement of embodied artificial intelligence. Current datasets and models for 3D visual grounding predominantly focus on identifying and loca…

3D visual groundingVisual Grounding

InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning

2025-05-23 · Zifu Wan, Yaqi Xie, Ce Zhang, Zhiqiu Lin 외

Large multimodal foundation models, particularly in the domains of language and vision, have significantly advanced various tasks, including robotics, autonomous driving, information retrieval, and grounding. However, ma…

Autonomous DrivingInformation RetrievalRetrievalSegmentation

OmniLayout: A Schematic-Coupled Multimodal Benchmark for Constraint-Aware Geometric Reasoning in PCB Layout

2026-07-03 · Taiting Lu, Kaiyuan Lin, Mingjia Wang, Haolin Ye 외 arxiv

Recent large language models (LLMs) have demonstrated remarkable progress in 3D spatial reasoning, spatial grounding, and fine-grained geometric understanding. However, their ability to reason about densely packed object…

Spatial Reasoning