paper-with-me

Papers

Large Language Model-empowered multimodal strain sensory system for shape recognition, monitoring, and human interaction of tensegrity

2024-06-11 · Zebing Mao, Ryota Kobayashi, Hiroyuki Nabae, Koichi Suzumori

A tensegrity-based system is a promising approach for dynamic exploration of uneven and unpredictable environments, particularly, space exploration. However, implementing such systems presents challenges in terms of intelligent aspects: state recognition, wireless monitoring, human interaction, and smart analyzing and advising function. Here, we introduce a 6-strut tensegrity integrate with 24 multimodal strain sensors by leveraging both deep learning model and large language models to realize smart tensegrity. Using conductive flexible tendons assisted by long short-term memory model, the tensegrity achieves the self-shape reconstruction without extern sensors. Through integrating the flask server and gpt-3.5-turbo model, the tensegrity autonomously enables to send data to iPhone for wireless monitoring and provides data analysis, explanation, prediction, and suggestions to human for decision making. Finally, human interaction system of the tensegrity helps human obtain necessary information of tensegrity from the aspect of human language. Overall, this intelligent tensegrity-based system with self-sensing tendons showcases potential for future exploration, making it a versatile tool for real-world applications.

📄 PDF Abstract BibTeX arXiv:2406.10264

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingLanguage ModelingLanguage ModellingLarge Language Model

Similar Papers 제목 키워드 기반

Chat-to-Design: AI Assisted Personalized Fashion Design

2022-07-03 · Weiming Zhuang, Chongjie Ye, Ying Xu, Pengzhi Mao 외

In this demo, we present Chat-to-Design, a new multimodal interaction system for personalized fashion design. Compared to classic systems that recommend apparel based on keywords, Chat-to-Design enables users to design c…

multimodal interactionNatural Language UnderstandingRetrieval

ENWAR: A RAG-empowered Multi-Modal LLM Framework for Wireless Environment Perception

2024-10-08 · Ahmad M. Nazar, Abdulkadir Celik, Mohamed Y. Selim, Asmaa Abdallah 외

Large language models (LLMs) hold significant promise in advancing network management and orchestration in 6G and beyond networks. However, existing LLMs are limited in domain-specific knowledge and their ability to hand…

RAGRetrieval-augmented Generation

Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition

2026-05-31 · Wanlong Fang, Tianle Zhang, Wen Tao, Alvin Chan arxiv

Understanding modality interaction in multimodal large language models (MLLMs) is central to reliable deployment. We introduce Partial Information Decomposition (PID) as a decision-level framework that separates unique, …

Multimodal Reasoning

Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings

2025-03-10 · Jonghyun Lee, Dojun Park, Jiwoo Lee, Hoekeon Choi 외

This study investigated the multimodal perception of large language models (LLMs), focusing on their ability to capture human-like perceptual strength ratings across sensory modalities. Utilizing perceptual strength rati…

MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic Manipulation

2025-09-30 · Zhuoyang Liu, Jiaming Liu, Jiadong Xu, Nuowei Han 외 arxiv

Vision-language-action models (VLAs) have shown generalization capabilities in robotic manipulation tasks by inheriting from vision-language models (VLMs) and learning action generation. Most VLA models focus on interpre…

Point Clouds