paper-with-me

홈 › Papers

Adaptive RAN Slicing Control via Reward-Free Self-Finetuning Agents

2026-03-11 · Yuanhao Li, Haozhe Wang, Geyong Min, Nektarios Georgalas, Wang Miao arxiv

The integration of Generative AI models into AI-native network systems offers a transformative path toward achieving autonomous and adaptive control. However, the application of such models to continuous control tasks is impeded by intrinsic architectural limitations, including finite context windows, the lack of explicit reward signals, and the degradation of the long context. This paper posits that the key to unlocking robust continuous control is enabling agents to internalize experience by distilling it into their parameters, rather than relying on prompt-based memory. To this end, we propose a novel self-finetuning framework that enables agentic systems to learn continuously through direct interaction with the environment, bypassing the need for handcrafted rewards. Our framework implements a bi-perspective reflection mechanism that generates autonomous linguistic feedback to construct preference datasets from interaction history. A subsequent preference-based fine-tuning process distills long-horizon experiences into the model's parameters. We evaluate our approach on a dynamic Radio Access Network (RAN) slicing task, a challenging multi-objective control problem that requires the resolution of acute trade-offs between spectrum efficiency, service quality, and reconfiguration stability under volatile network conditions. Experimental results show that our framework outperforms standard Reinforcement Learning (RL) baselines and existing Large Language Model (LLM)-based agents in sample efficiency, stability, and multi-metric optimization. These findings demonstrate the potential of self-improving generative agents for continuous control tasks, paving the way for future AI-native network infrastructure.

📄 PDF Abstract BibTeX arXiv:2603.10564

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningContinuous Control

Similar Papers 제목 키워드 기반

Adaptive Resource Management for Edge Network Slicing using Incremental Multi-Agent Deep Reinforcement Learning

2023-10-26 · Haiyuan Li, Yuelin Liu, Xueqing Zhou, Xenofon Vasilakos 외

Multi-access edge computing provides local resources in mobile networks as the essential means for meeting the demands of emerging ultra-reliable low-latency communications. At the edge, dynamic computing requests requir…

Deep Reinforcement LearningEdge-computingIncremental LearningManagement+1

Deep Reinforcement Learning for Adaptive Network Slicing in 5G for Intelligent Vehicular Systems and Smart Cities

2020-10-19 · Almuthanna Nassar, Yasin Yilmaz

Intelligent vehicular systems and smart city applications are the fastest growing Internet of things (IoT) implementations at a compound annual growth rate of 30%. In view of the recent advances in IoT devices and the em…

Deep Reinforcement Learning

Adaptive remanufacturing for freeform surface parts based on linear laser scanner and robotic laser cladding

2024-08-10 · Robotics and Computer-Integrated Manufacturing 2024 8 · Wei Ma, TianliangHu, Chengrui Zhang, Qizhi Chen

Freeform surface parts play a significant role in the aerospace industry, the mold- manufacturing industry and the automobile industry, and it is energy-saving, material-saving, time-saving and environmentally beneficial…

Relation-Aware Slicing in Cross-Domain Alignment

2025-07-17 · Dhruv Sarkar, Aprameyo Chakrabartty, Anish Chakrabarty, Swagatam Das arxiv

The Sliced Gromov-Wasserstein (SGW) distance, aiming to relieve the computational cost of solving a non-convex quadratic program that is the Gromov-Wasserstein distance, utilizes projecting directions sampled uniformly f…

Adversarial Machine Learning for Flooding Attacks on 5G Radio Access Network Slicing

2021-01-21 · Yi Shi, Yalin E. Sagduyu

Network slicing manages network resources as virtual resource blocks (RBs) for the 5G Radio Access Network (RAN). Each communication request comes with quality of experience (QoE) requirements such as throughput and late…

BIG-bench Machine LearningReinforcement Learning (RL)