Beyond Interpretability: Exploring the Comprehensibility of Adaptive Video Streaming through Large Language Models
Over the past decade, adaptive video streaming technology has witnessed significant advancements, particularly driven by the rapid evolution of deep learning techniques. However, the black-box nature of deep learning algorithms presents challenges for developers in understanding decision-making processes and optimizing for specific application scenarios. Although existing research has enhanced algorithm interpretability through decision tree conversion, interpretability does not directly equate to developers' subjective comprehensibility. To address this challenge, we introduce \texttt{ComTree}, the first bitrate adaptation algorithm generation framework that considers comprehensibility. The framework initially generates the complete set of decision trees that meet performance requirements, then leverages large language models to evaluate these trees for developer comprehensibility, ultimately selecting solutions that best facilitate human understanding and enhancement. Experimental results demonstrate that \texttt{ComTree} significantly improves comprehensibility while maintaining competitive performance, showing potential for further advancement. The source code is available at https://github.com/thu-media/ComTree.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
A review of possible effects of cognitive biases on the interpretation of rule-based machine learning models
While the interpretability of machine learning models is often equated with their mere syntactic comprehensibility, we think that interpretability goes beyond that, and that human interpretability should also be investig…
BIG-bench Machine LearningInterpretable Machine LearningSubjective Evaluation of Comprehensibility in Movie Interactions
Various research works have dealt with the comprehensibility of textual, audio, or audiovisual documents, and showed that factors related to text (e.g. linguistic complexity), sound (e.g. speech intelligibility), image (…
Discriminatory Expressions to Produce Interpretable Models in Short Documents
Social Networking Sites (SNS) are one of the most important ways of communication. In particular, microblogging sites are being used as analysis avenues due to their peculiarities (promptness, short texts...). There are …
feature selectionBeyond Interpretability: The Gains of Feature Monosemanticity on Model Robustness
Deep learning models often suffer from a lack of interpretability due to polysemanticity, where individual neurons are activated by multiple unrelated semantics, resulting in unclear attributions of model behavior. Recen…
Domain GeneralizationFew-Shot LearningAdaptive Semantic-Spatio-Temporal Graph Convolutional Network for Lip Reading
The goal of this work is to recognize words, phrases, and sentences being spoken by a talking face without given the audio. Current deep learning approaches for lip reading focus on exploring the appearance and optical f…
Landmark-based LipreadingLip ReadingOptical Flow Estimation