paper-with-me

Papers

Network Aware Compute and Memory Allocation in Optically Composable Data Centres with Deep Reinforcement Learning and Graph Neural Networks

2022-10-26 · Zacharaya Shabka, Georgios Zervas

Resource-disaggregated data centre architectures promise a means of pooling resources remotely within data centres, allowing for both more flexibility and resource efficiency underlying the increasingly important infrastructure-as-a-service business. This can be accomplished by means of using an optically circuit switched backbone in the data centre network (DCN); providing the required bandwidth and latency guarantees to ensure reliable performance when applications are run across non-local resource pools. However, resource allocation in this scenario requires both server-level \emph{and} network-level resource to be co-allocated to requests. The online nature and underlying combinatorial complexity of this problem, alongside the typical scale of DCN topologies, makes exact solutions impossible and heuristic based solutions sub-optimal or non-intuitive to design. We demonstrate that \emph{deep reinforcement learning}, where the policy is modelled by a \emph{graph neural network} can be used to learn effective \emph{network-aware} and \emph{topologically-scalable} allocation policies end-to-end. Compared to state-of-the-art heuristics for network-aware resource allocation, the method achieves up to $20\%$ higher acceptance ratio; can achieve the same acceptance ratio as the best performing heuristic with $3\times$ less networking resources available and can maintain all-around performance when directly applied (with no further training) to DCN topologies with $10^2\times$ more servers than the topologies seen during training.

📄 PDF Abstract BibTeX arXiv:2211.02466

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement LearningGraph Neural Network

Similar Papers 제목 키워드 기반

Optically Writable Atomic Vapor Memory as a Substrate for Optical Reservoir Computing

2026-08-18 · Elizabeth Robertson, Mingwei Yang, Lina Jaurigue, Guillermo Gallego 외 arxiv

We present an optical random access memory (ORAM) based on warm cesium (Cs) atomic vapor and demonstrate its operation as the physical substrate of a reservoir computer. Information is stored in the hyperfine population …

SALT: Salience-Aware Lexical Trie for Long-Context Compression

2026-07-20 · Oteo Mamo, Hyunjin Yi, Joydhriti Choudhury, Shangqian Gao 외 arxiv

As large language models (LLMs) process increasingly longer prompts, computation and KV-cache memory costs have emerged as major bottlenecks in inference systems. Existing input-level prompt compression methods address t…

Unifying Data, Memory, and Compute Efficiency in LLM training: A Survey

2026-06-09 · Vanessa Schmidt, Huy Hoang Nguyen, Cédric Jung, Shirin Salehi 외 arxiv

Resource constraints increasingly determine what can be trained, fine-tuned, and deployed in large language models (LLMs), yet efficiency is often studied through isolated techniques rather than as an interacting system …

Second-Order, First-Class: A Composable Stack for Curvature-Aware Training

2026-03-26 · Mikalai Korbit, Mario Zanon arxiv

Second-order methods promise improved stability and faster convergence, yet they remain underused due to implementation overhead, tuning brittleness, and the lack of composable APIs. We introduce Somax, a composable Opta…

KRISM --- Krylov Subspace-based Optical Computing of Hyperspectral Images

2018-01-26 · Vishwanath Saragadam, Aswin C. Sankaranarayanan

We present an adaptive imaging technique that optically computes a low-rank approximation of a scene's hyperspectral image, conceptualized as a matrix. Central to the proposed technique is the optical implementation of t…