paper-with-me

Papers

LangGrasp: Leveraging Fine-Tuned LLMs for Language Interactive Robot Grasping with Ambiguous Instructions

2025-10-02 · Yunhan Lin, Wenqi Wu, Zhijie Zhang, Huasong Min arxiv

The existing language-driven grasping methods struggle to fully handle ambiguous instructions containing implicit intents. To tackle this challenge, we propose LangGrasp, a novel language-interactive robotic grasping framework. The framework integrates fine-tuned large language models (LLMs) to leverage their robust commonsense understanding and environmental perception capabilities, thereby deducing implicit intents from linguistic instructions and clarifying task requirements along with target manipulation objects. Furthermore, our designed point cloud localization module, guided by 2D part segmentation, enables partial point cloud localization in scenes, thereby extending grasping operations from coarse-grained object-level to fine-grained part-level manipulation. Experimental results show that the LangGrasp framework accurately resolves implicit intents in ambiguous instructions, identifying critical operations and target information that are unstated yet essential for task completion. Additionally, it dynamically selects optimal grasping poses by integrating environmental information. This enables high-precision grasping from object-level to part-level manipulation, significantly enhancing the adaptability and task execution efficiency of robots in unstructured environments. More information and code are available here: https://github.com/wu467/LangGrasp.

📄 PDF Abstract BibTeX arXiv:2510.02104

Code (0)

등록된 구현이 없습니다.

Tasks

Robotic Grasping

Similar Papers 제목 키워드 기반

Augmented Relevance Datasets with Fine-Tuned Small LLMs

2025-04-14 · Quentin Fitte-Rey, Matyas Amrouche, Romain Deveaud

Building high-quality datasets and labeling query-document relevance are essential yet resource-intensive tasks, requiring detailed guidelines and substantial effort from human annotators. This paper explores the use of …

Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing

2024-06-10 · Enshuo Hsu, Kirk Roberts

The performance of deep learning-based natural language processing systems is based on large amounts of labeled training data which, in the clinical domain, are not easily available or affordable. Weak supervision and in…

In-Context Learning

Unlocking the Potential of User Feedback: Leveraging Large Language Model as User Simulator to Enhance Dialogue System

2023-06-16 · Zhiyuan Hu, Yue Feng, Anh Tuan Luu, Bryan Hooi 외

Dialogue systems and large language models (LLMs) have gained considerable attention. However, the direct utilization of LLMs as task-oriented dialogue (TOD) models has been found to underperform compared to smaller task…

Language ModelingLanguage ModellingLarge Language Model

PRISMA-DFLLM: An Extension of PRISMA for Systematic Literature Reviews using Domain-specific Finetuned Large Language Models

2023-06-15 · Teo Susnjak

With the proliferation of open-sourced Large Language Models (LLMs) and efficient finetuning techniques, we are on the cusp of the emergence of numerous domain-specific LLMs that have been finetuned for expertise across …

Taiyi: A Bilingual Fine-Tuned Large Language Model for Diverse Biomedical Tasks

2023-11-20 · Ling Luo, Jinzhong Ning, Yingwen Zhao, Zhijun Wang 외

Objective: Most existing fine-tuned biomedical large language models (LLMs) focus on enhancing performance in monolingual biomedical question answering and conversation tasks. To investigate the effectiveness of the fine…

Language ModelingLanguage ModellingLarge Language Modelnamed-entity-recognition+5