paper-with-me

홈 › Papers

LILA: Language-Informed Latent Actions

2021-11-05 · Siddharth Karamcheti, Megha Srivastava, Percy Liang, Dorsa Sadigh

We introduce Language-Informed Latent Actions (LILA), a framework for learning natural language interfaces in the context of human-robot collaboration. LILA falls under the shared autonomy paradigm: in addition to providing discrete language inputs, humans are given a low-dimensional controller $-$ e.g., a 2 degree-of-freedom (DoF) joystick that can move left/right and up/down $-$ for operating the robot. LILA learns to use language to modulate this controller, providing users with a language-informed control space: given an instruction like "place the cereal bowl on the tray," LILA may learn a 2-DoF space where one dimension controls the distance from the robot's end-effector to the bowl, and the other dimension controls the robot's end-effector pose relative to the grasp point on the bowl. We evaluate LILA with real-world user studies, where users can provide a language instruction while operating a 7-DoF Franka Emika Panda Arm to complete a series of complex manipulation tasks. We show that LILA models are not only more sample efficient and performant than imitation learning and end-effector control baselines, but that they are also qualitatively preferred by users.

📄 PDF Abstract BibTeX arXiv:2111.03205

Code (1)

siddk/lila 공식 구현 pytorch

Tasks

Imitation Learning

Similar Papers 제목 키워드 기반

"No, to the Right" -- Online Language Corrections for Robotic Manipulation via Shared Autonomy

2023-01-06 · Yuchen Cui, Siddharth Karamcheti, Raj Palleti, Nidhya Shivakumar 외

Systems for language-guided human-robot interaction must satisfy two key desiderata for broad adoption: adaptivity and learning efficiency. Unfortunately, existing instruction-following agents cannot adapt, lacking the a…

Instruction Following

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

2026-08-04 · Fan Yang, Yuting Su, Xiaobo Wang, Yuncheng You 외 arxiv

World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticipate how a scene will evolve. However, existing WAMs often incur substa…

DeliLaw: A Chinese Legal Counselling System Based on a Large Language Model

2024-08-01 · Nan Xie, Yuelin Bai, Hengyuan Gao, Feiteng Fang 외

Traditional legal retrieval systems designed to retrieve legal documents, statutes, precedents, and other legal information are unable to give satisfactory answers due to lack of semantic understanding of specific questi…

ArticlesHallucinationLanguage ModelingLanguage Modelling+2

LILAC: Language-Conditioned Object-Centric Optical Flow for Open-Loop Trajectory Generation

2026-03-26 · Motonari Kambara, Koki Seno, Tomoya Kaichi, Yanan Wang 외 arxiv

We address language-conditioned robotic manipulation using flow-based trajectory generation, which enables training on human and web videos of object manipulation and requires only minimal embodiment-specific data. This …

LiLa-Net: Lightweight Latent LiDAR Autoencoder for 3D Point Cloud Reconstruction

2025-10-02 · Mario Resino, Borja Pérez, Jaime Godoy, Abdulla Al-Kaff 외 arxiv

This work proposed a 3D autoencoder architecture, named LiLa-Net, which encodes efficient features from real traffic environments, employing only the LiDAR's point clouds. For this purpose, we have real semi-autonomous v…

Point Clouds