paper-with-me

홈 › Papers

Hybrid Reasoning for Perception, Explanation, and Autonomous Action in Manufacturing

2025-06-10 · Christos Margadji, Sebastian W. Pattinson

Industrial processes must be robust and adaptable, as environments and tasks are often unpredictable, while operational errors remain costly and difficult to detect. AI-based control systems offer a path forward, yet typically depend on supervised learning with extensive labelled datasets, which limits their ability to generalize across variable and data-scarce industrial settings. Foundation models could enable broader reasoning and knowledge integration, but rarely deliver the quantitative precision demanded by engineering applications. Here, we introduceControl and Interpretation of Production via Hybrid Expertise and Reasoning (CIPHER): a vision-language-action (VLA) model framework aiming to replicate human-like reasoning for industrial control, instantiated in a commercial-grade 3D printer. It integrates a process expert, a regression model enabling quantitative characterization of system states required for engineering tasks. CIPHER also incorporates retrieval-augmented generation to access external expert knowledge and support physics-informed, chain-of-thought reasoning. This hybrid architecture exhibits strong generalization to out-of-distribution tasks. It interprets visual or textual inputs from process monitoring, explains its decisions, and autonomously generates precise machine instructions, without requiring explicit annotations. CIPHER thus lays the foundations for autonomous systems that act with precision, reason with context, and communicate decisions transparently, supporting safe and trusted deployment in industrial settings.

📄 PDF Abstract BibTeX arXiv:2506.08462

Code (0)

등록된 구현이 없습니다.

Tasks

Retrieval-augmented GenerationVision-Language-Action

Similar Papers 제목 키워드 기반

DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking

2025-07-28 · Weicheng Zheng, Xiaofei Mao, Nanfei Ye, Pengxiang Li 외 arxiv

The advent of Vision-Language Models (VLMs) has significantly advanced end-to-end autonomous driving, demonstrating powerful reasoning abilities for high-level behavior planning tasks. However, existing methods are often…

Reinforcement LearningAutonomous DrivingVisual Reasoning

RT-VLA: Real-Time Vision-Language-Action Models via Knowledge Distillation

2026-06-12 · Xiangyu Huang, Zhenlin Hua, Han Zhou, Shounak Sural 외 arxiv

Vision-Language-Action (VLA) models have shown strong potential for end-to-end autonomous driving by jointly modeling visual perception, language reasoning, explainability and action prediction. However, their large visi…

Knowledge DistillationAutonomous Driving

HyPerNav: Hybrid Perception for Object-Oriented Navigation in Unknown Environment

2025-10-27 · Zecheng Yin, Hao Zhao, Zhen Li arxiv

Objective-oriented navigation(ObjNav) enables robot to navigate to target object directly and autonomously in an unknown environment. Effective perception in navigation in unknown environment is critical for autonomous r…

RCA: Region Conditioned Adaptation for Visual Abductive Reasoning

2023-03-18 · Hao Zhang, Yeo Keat Ee, Basura Fernando

Visual abductive reasoning aims to make likely explanations for visual observations. We propose a simple yet effective Region Conditioned Adaptation, a hybrid parameter-efficient fine-tuning method that equips the frozen…

parameter-efficient fine-tuningVisual Abductive Reasoning

Simulation-based Scenario Generation for Robust Hybrid AI for Autonomy

2024-09-10 · Hambisa Keno, Nicholas J. Pioch, Christopher Guagliano, Timothy H. Chung

Application of Unmanned Aerial Vehicles (UAVs) in search and rescue, emergency management, and law enforcement has gained traction with the advent of low-cost platforms and sensor payloads. The emergence of hybrid neural…