paper-with-me

홈 › Papers

TeNet: Text-to-Network for Compact Policy Synthesis

2026-01-22 · Ariyan Bighashdel, Kevin Sebastian Luck arxiv

Robots that follow natural-language instructions often either plan at a high level using hand-designed interfaces or rely on large end-to-end models that are difficult to deploy for real-time control. We propose TeNet (Text-to-Network), a framework for instantiating compact, task-specific robot policies directly from natural language descriptions. TeNet conditions a hypernetwork on text embeddings produced by a pretrained large language model (LLM) to generate a fully executable policy, which then operates solely on low-dimensional state inputs at high control frequencies. By using the language only once at the policy instantiation time, TeNet inherits the general knowledge and paraphrasing robustness of pretrained LLMs while remaining lightweight and efficient at execution time. To improve generalization, we optionally ground language in behavior during training by aligning text embeddings with demonstrated actions, while requiring no demonstrations at inference time. Experiments on MuJoCo and Meta-World benchmarks show that TeNet produces policies that are orders of magnitude smaller than sequence-based baselines, while achieving strong performance in both multi-task and meta-learning settings and supporting high-frequency control. These results show that text-conditioned hypernetworks offer a practical way to build compact, language-driven controllers for ressource-constrained robot control tasks with real-time requirements.

📄 PDF Abstract BibTeX arXiv:2601.15912

Code (0)

등록된 구현이 없습니다.

Tasks

General Knowledge

Similar Papers 제목 키워드 기반

TwinLiteNetPlus: A Stronger Model for Real-time Drivable Area and Lane Segmentation

2024-03-25 · Quang-Huy Che, Duc-Tri Le, Minh-Quan Pham, Vinh-Tiep Nguyen 외

Semantic segmentation is crucial for autonomous driving, particularly for Drivable Area and Lane Segmentation, ensuring safety and navigation. To address the high computational costs of current state-of-the-art (SOTA) mo…

Autonomous DrivingDrivable Area DetectionLane DetectionSegmentation+1

Joint Reference Frame Synthesis and Post Filter Enhancement for Versatile Video Coding

2024-04-28 · Weijie Bao, Yuantong Zhang, Jianghao Jia, Zhenzhong Chen 외

This paper presents the joint reference frame synthesis (RFS) and post-processing filter enhancement (PFE) for Versatile Video Coding (VVC), aiming to explore the combination of different neural network-based video codin…

TENET: One Step Toward Test-Driven Development for Repository-Level Code Generation

2025-09-29 · Yiran Hu, Nan Jiang, Shanchao Liang, Yi Wu 외 arxiv

Test-Driven Development (TDD) is a widely adopted practice that requires developers to create and execute tests alongside implementation. With recent advances in Large Language Models (LLMs), developers can shift from ma…

Code Generation

RewriteNet: Reliable Scene Text Editing with Implicit Decomposition of Text Contents and Styles

2021-07-23 · Junyeop Lee, Yoonsik Kim, Seonghyeon Kim, Moonbin Yim 외

Scene text editing (STE), which converts a text in a scene image into the desired text while preserving an original style, is a challenging task due to a complex intervention between text and style. In this paper, we pro…

Image GenerationScene Text EditingScene Text Recognition

Learning the Model Update for Siamese Trackers

2019-08-02 · ICCV 2019 10 · Lichao Zhang, Abel Gonzalez-Garcia, Joost Van de Weijer, Martin Danelljan 외

Siamese approaches address the visual tracking problem by extracting an appearance template from the current frame, which is used to localize the target in the next frame. In general, this template is linearly combined w…

modelVisual Tracking