paper-with-me

Papers

Low Complexity Adaptive Machine Learning Approaches for End-to-End Latency Prediction

2023-01-31 · Pierre Larrenie, Jean-François Bercher, Olivier Venard, Iyad Lahsen-Cherif

Software Defined Networks have opened the door to statistical and AI-based techniques to improve efficiency of networking. Especially to ensure a certain Quality of Service (QoS) for specific applications by routing packets with awareness on content nature (VoIP, video, files, etc.) and its needs (latency, bandwidth, etc.) to use efficiently resources of a network. Monitoring and predicting various Key Performance Indicators (KPIs) at any level may handle such problems while preserving network bandwidth. The question addressed in this work is the design of efficient, low-cost adaptive algorithms for KPI estimation, monitoring and prediction. We focus on end-to-end latency prediction, for which we illustrate our approaches and results on data obtained from a public generator provided after the recent international challenge on GNN [12]. In this paper, we improve our previously proposed low-cost estimators [6] by adding the adaptive dimension, and show that the performances are minimally modified while gaining the ability to track varying networks.

📄 PDF Abstract BibTeX arXiv:2301.13536

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Golden Queue Managers 설명 없음

Similar Papers 제목 키워드 기반

Data-Driven Adaptive Simultaneous Machine Translation

2022-04-27 · Guangxu Xun, Mingbo Ma, Yuchen Bian, Xingyu Cai 외

In simultaneous translation (SimulMT), the most widely used strategy is the wait-k policy thanks to its simplicity and effectiveness in balancing translation quality and latency. However, wait-k suffers from two major li…

Machine TranslationSentenceTranslation

Data-Driven Adaptive Simultaneous Machine Translation

2021-11-16 · ACL ARR November 2021 11 · Anonymous

In simultaneous translation (SimulMT), the most widely used strategy is the \waitk policy thanks to its simplicity and effectiveness in balancing translation quality and latency. However, \waitk suffers from two major li…

Machine TranslationSentenceTranslation

CacheNet: A Model Caching Framework for Deep Learning Inference on the Edge

2020-07-03 · Yihao Fang, Shervin Manzuri Shalmani, Rong Zheng

The success of deep neural networks (DNN) in machine perception applications such as image classification and speech recognition comes at the cost of high computation and storage complexity. Inference of uncompressed lar…

image-classificationImage Classificationspeech-recognitionSpeech Recognition

Clipper: A Low-Latency Online Prediction Serving System

2016-12-09 · Daniel Crankshaw, Xin Wang, Giulio Zhou, Michael J. Franklin 외

Machine learning is being deployed in a growing number of applications which demand real-time, accurate, and robust predictions under heavy query load. However, most machine learning frameworks and systems only address m…

BIG-bench Machine LearningModel SelectionPrediction

Adaptive ToR: Complexity-Aware Tree-Based Retrieval for Pareto-Optimal Multi-Intent NLU

2026-04-27 · Hee-Kyong Yoo, Wonbae Kim, Hyocheol Ahn arxiv

Multi-intent natural language understanding requires retrieval systems that simultaneously achieve high accuracy and computational efficiency, yet existing approaches apply either uniform single-step retrieval that compr…

Natural Language UnderstandingComputational Efficiency