paper-with-me

홈 › Papers

Human-in-the-Loop Bandwidth Estimation for Quality of Experience Optimization in Real-Time Video Communication

2025-10-14 · Sami Khairy, Gabriel Mittag, Vishak Gopal, Ross Cutler arxiv

The quality of experience (QoE) delivered by video conferencing systems is significantly influenced by accurately estimating the time-varying available bandwidth between the sender and receiver. Bandwidth estimation for real-time communications remains an open challenge due to rapidly evolving network architectures, increasingly complex protocol stacks, and the difficulty of defining QoE metrics that reliably improve user experience. In this work, we propose a deployed, human-in-the-loop, data-driven framework for bandwidth estimation to address these challenges. Our approach begins with training objective QoE reward models derived from subjective user evaluations to measure audio and video quality in real-time video conferencing systems. Subsequently, we collect roughly $1$M network traces with objective QoE rewards from real-world Microsoft Teams calls to curate a bandwidth estimation training dataset. We then introduce a novel distributional offline reinforcement learning (RL) algorithm to train a neural-network-based bandwidth estimator aimed at improving QoE for users. Our real-world A/B test demonstrates that the proposed approach reduces the subjective poor call ratio by $11.41\%$ compared to the baseline bandwidth estimator. Furthermore, the proposed offline RL algorithm is benchmarked on D4RL tasks to demonstrate its generalization beyond bandwidth estimation.

📄 PDF Abstract BibTeX arXiv:2510.12265

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningOffline RL

Similar Papers 제목 키워드 기반

Towards a control-theoretic trivialization of ABR video streaming

2023-10-02 · Michel Fliess, Cédric Join

Combining flatness-based control, model-free control and algebraic estimation techniques permits to trivialize several key issues in the adaptive bitrate (ABR) setting, now the dominant industry approach in video streami…

EOM Minimum Point Bias Voltage Estimation for Application in Quantum Computing

2024-12-18 · Frank Obernosterer, Raimund Meyer, Robert Koch, Gerd Kilian 외

In quantum computing systems the quantum states of qubits can be modified among others by applying light pulses. In order to achieve low computing error rates these pulses have to be precisely shaped in magnitude and pha…

Robust Bandwidth Estimation for Real-Time Communication with Offline Reinforcement Learning

2025-07-08 · Jian Kai, Tianwei Zhang, Zihan Ling, Yang Cao 외

Accurate bandwidth estimation (BWE) is critical for real-time communication (RTC) systems. Traditional heuristic approaches offer limited adaptability under dynamic networks, while online reinforcement learning (RL) suff…

Offline RLReinforcement Learning (RL)

Reinforcement learning for bandwidth estimation and congestion control in real-time communications

2019-12-04 · Joyce Fang, Martin Ellis, Bin Li, Siyao Liu 외

Bandwidth estimation and congestion control for real-time communications (i.e., audio and video conferencing) remains a difficult problem, despite many years of research. Achieving high quality of experience (QoE) for en…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

ROVE: Unlocking Human Interventions for Humanoid Manipulation via Reinforcement Learning

2026-06-15 · Wei Xiao, Weiliang Tang, Yuying Ge, Hui Zhou 외 arxiv

Human interventions provide crucial corrective signals for post-training Vision-Language-Action (VLA) models. However, enabling seamless humanoid interventions is a formidable systems challenge due to complex whole-body …

Reinforcement Learning