paper-with-me

Papers

Probability Hacking and the Design of Trustworthy ML for Signal Processing in C-UAS: A Scenario Based Method

2026-02-08 · Liisa Janssens, Laura Middeldorp arxiv

In order to counter the various threats manifested by Unmanned Aircraft Systems (UAS) adequately, specialized Counter Unmanned Aircraft Systems (C-UAS) are required. Enhancing C-UAS with Emerging and Disruptive Technologies (EDTs) such as Artificial Intelligence (AI) can lead to more effective countermeasures. In this paper a scenario-based method is applied to C-UAS augmented with Machine Learning (ML), a subset of AI, that can enhance signal processing capabilities. Via the scenarios-based method we frame in this paper probability hacking as a challenge and identify requirements which can be implemented in existing Rule of Law mechanisms to prevent probability hacking. These requirements strengthen the trustworthiness of the C-UAS, which feed into justified trust - a key to successful Human-Autonomy Teaming, in civil and military contexts. Index Terms: C-UAS, Scenario-based method, Emerging and Disruptive Technologies, Probability hacking, Trustworthiness.

📄 PDF Abstract BibTeX arXiv:2602.08086

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

From Passive Metric to Active Signal: The Evolving Role of Uncertainty Quantification in Large Language Models

2026-01-22 · Jiaxin Zhang, Wendi Cui, Zhuohang Li, Lifu Huang 외 arxiv

While Large Language Models (LLMs) show remarkable capabilities, their unreliability remains a critical barrier to deployment in high-stakes domains. This survey charts a functional evolution in addressing this challenge…

Reinforcement Learning

Sail into the Headwind: Alignment via Robust Rewards and Dynamic Labels against Reward Hacking

2024-12-12 · Paria Rashidinejad, Yuandong Tian

Aligning AI systems with human preferences typically suffers from the infamous reward hacking problem, where optimization of an imperfect reward model leads to undesired behaviors. In this paper, we investigate reward ha…

Mathematical Reasoning

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

2026-05-20 · Amit Roth, Ankur Samanta, Matan Halevy, Yoav Levine 외 arxiv

Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby agents appear successful under the evaluation signal while violating…

What Makes a Medical Checker Trainable? Diagnosing Signal Collapse and Reward Hacking in Checker-Guided RAG for Biomedical QA

2026-05-25 · Yuelyu Ji, Min Gu Kwak, Hang Zhang, Xizhi Wu 외 arxiv

Medical RAG needs evidence-grounded claims, so plugging a claim-level NLI checker into retrieval-augmented RL is intuitive. \textbf{We find that the checker's \emph{output distribution} during training, not its held-out …

Conformal Prediction for Trustworthy Detection of Railway Signals

2023-01-26 · Léo Andéol, Thomas Fel, Florence De Grancey, Luca Mossina

We present an application of conformal prediction, a form of uncertainty quantification with guarantees, to the detection of railway signals. State-of-the-art architectures are tested and the most promising one undergoes…

Conformal PredictionPredictionUncertainty Quantification