paper-with-me

홈 › Papers

Canaries and Whistles: Resilient Drone Communication Networks with (or without) Deep Reinforcement Learning

2023-12-08 · Chris Hicks, Vasilios Mavroudis, Myles Foley, Thomas Davies, Kate Highnam, Tim Watson

Communication networks able to withstand hostile environments are critically important for disaster relief operations. In this paper, we consider a challenging scenario where drones have been compromised in the supply chain, during their manufacture, and harbour malicious software capable of wide-ranging and infectious disruption. We investigate multi-agent deep reinforcement learning as a tool for learning defensive strategies that maximise communications bandwidth despite continual adversarial interference. Using a public challenge for learning network resilience strategies, we propose a state-of-the-art expert technique and study its superiority over deep reinforcement learning agents. Correspondingly, we identify three specific methods for improving the performance of our learning-based agents: (1) ensuring each observation contains the necessary information, (2) using expert agents to provide a curriculum for learning, and (3) paying close attention to reward. We apply our methods and present a new mixed strategy enabling expert and learning-based agents to work together and improve on all prior results.

📄 PDF Abstract BibTeX arXiv:2312.04940

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learningreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Resilient Distribution System Restoration with Communication Recovery by Drone Small Cells

2022-03-31 · Haochen Zhang, Chen Chen, Shunbo Lei, Zhaohong Bie

Distribution system (DS) restoration after natural disasters often faces the challenge of communication failures to feeder automation (FA) facilities, resulting in prolonged load pick-up process. This letter discusses th…

Back to square one: probabilistic trajectory forecasting without bells and whistles

2018-12-07 · Ehsan Pajouheshgar, Christoph H. Lampert

We introduce a spatio-temporal convolutional neural network model for trajectory forecasting from visual sources. Applied in an auto-regressive way it provides an explicit probability distribution over continuations of a…

RelationTrajectory Forecasting

Deep Q-Network Based Resilient Drone Communication:Neutralizing First-Order Markov Jammers

2026-01-01 · Andrii Grekhov, Volodymyr Kharchenko, Vasyl Kondratiuk arxiv

Deep Reinforcement Learning based solution for jamming communications using Frequency Hopping Spread Spectrum technology in a 16 channel radio environment is presented. Deep Q Network based transmitter continuously selec…

Reinforcement Learning

Unleashing the Power of Randomization in Auditing Differentially Private ML

2023-05-29 · NeurIPS 2023 11

We present a rigorous methodology for auditing differentially private machine learning algorithms by adding multiple carefully designed examples called canaries. We take a first principles approach based on three key com…

Advancing the State-of-the-Art in Empirical Privacy Auditing

2026-06-09 · Nicole Mitchell, Galen Andrew, Arun Ganesh, Brendan McMahan 외 arxiv

Parameter-efficient fine-tuning of large language models (LLMs) can exhibit problematic memorization of individual training examples. Empirical privacy auditing (EPA) quantifies this risk by measuring realistic data leak…

parameter-efficient fine-tuning