paper-with-me

홈 › Papers

Pretrained LLMs as Real-Time Controllers for Robot Operated Serial Production Line

2025-03-05 · Muhammad Waseem, Kshitij Bhatta, Chen Li, Qing Chang

The manufacturing industry is undergoing a transformative shift, driven by cutting-edge technologies like 5G, AI, and cloud computing. Despite these advancements, effective system control, which is crucial for optimizing production efficiency, remains a complex challenge due to the intricate, knowledge-dependent nature of manufacturing processes and the reliance on domain-specific expertise. Conventional control methods often demand heavy customization, considerable computational resources, and lack transparency in decision-making. In this work, we investigate the feasibility of using Large Language Models (LLMs), particularly GPT-4, as a straightforward, adaptable solution for controlling manufacturing systems, specifically, mobile robot scheduling. We introduce an LLM-based control framework to assign mobile robots to different machines in robot assisted serial production lines, evaluating its performance in terms of system throughput. Our proposed framework outperforms traditional scheduling approaches such as First-Come-First-Served (FCFS), Shortest Processing Time (SPT), and Longest Processing Time (LPT). While it achieves performance that is on par with state-of-the-art methods like Multi-Agent Reinforcement Learning (MARL), it offers a distinct advantage by delivering comparable throughput without the need for extensive retraining. These results suggest that the proposed LLM-based solution is well-suited for scenarios where technical expertise, computational resources, and financial investment are limited, while decision transparency and system scalability are critical concerns.

📄 PDF Abstract BibTeX arXiv:2503.03889

Code (0)

등록된 구현이 없습니다.

Tasks

Cloud ComputingMulti-agent Reinforcement LearningScheduling

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…

Similar Papers 제목 키워드 기반

TeNet: Text-to-Network for Compact Policy Synthesis

2026-01-22 · Ariyan Bighashdel, Kevin Sebastian Luck arxiv

Robots that follow natural-language instructions often either plan at a high level using hand-designed interfaces or rely on large end-to-end models that are difficult to deploy for real-time control. We propose TeNet (T…

General Knowledge

LaGO: Latent Action Guidance for Online Reinforcement Learning

2026-06-23 · Kuan-Yen Liu, Ren-Jyun Huang, Ti-Rong Wu arxiv

Large language models (LLMs) have shown strong potential for planning and sequential decision-making, but prior work often relies on using them as direct controllers, which requires precise action generation and can be u…

Reinforcement Learning

Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception

2023-08-10 · Yunhao Yang, Cyrus Neary, Ufuk Topcu

Recently developed pretrained models can encode rich world knowledge expressed in multiple modalities, such as text and images. However, the outputs of these models cannot be integrated into algorithms to solve sequentia…

Decision MakingRobot ManipulationSequential Decision MakingWorld Knowledge

Efficient Learning of Control Policies for Robust Quadruped Bounding using Pretrained Neural Networks

2020-11-01 · Zhicheng Wang, Anqiao Li, Yixiao Zheng, Anhuan Xie 외

Bounding is one of the important gaits in quadrupedal locomotion for negotiating obstacles. The authors proposed an effective approach that can learn robust bounding gaits more efficiently despite its large variation in …

Deep Reinforcement LearningFeature Engineering

Double Q-PID algorithm for mobile robot control

2018-11-01 · IgnacioCarlucho, MarianoDe Paula, Gerardo G.Acosta

Many expert systems have been developed for self-adaptive PID controllers of mobile robots. However, the high computational requirements of the expert systems layers, developed for the tuning of the PID controllers, stil…

Active LearningQ-Learning