paper-with-me

홈 › Papers

DeepProv: Behavioral Characterization and Repair of Neural Networks via Inference Provenance Graph Analysis

2025-09-30 · Firas Ben Hmida, Abderrahmen Amich, Ata Kaboudi, Birhanu Eshete arxiv

Deep neural networks (DNNs) are increasingly being deployed in high-stakes applications, from self-driving cars to biometric authentication. However, their unpredictable and unreliable behaviors in real-world settings require new approaches to characterize and ensure their reliability. This paper introduces DeepProv, a novel and customizable system designed to capture and characterize the runtime behavior of DNNs during inference by using their underlying graph structure. Inspired by system audit provenance graphs, DeepProv models the computational information flow of a DNN's inference process through Inference Provenance Graphs (IPGs). These graphs provide a detailed structural representation of the behavior of DNN, allowing both empirical and structural analysis. DeepProv uses these insights to systematically repair DNNs for specific objectives, such as improving robustness, privacy, or fairness. We instantiate DeepProv with adversarial robustness as the goal of model repair and conduct extensive case studies to evaluate its effectiveness. Our results demonstrate its effectiveness and scalability across diverse classification tasks, attack scenarios, and model complexities. DeepProv automatically identifies repair actions at the node and edge-level within IPGs, significantly enhancing the robustness of the model. In particular, applying DeepProv repair strategies to just a single layer of a DNN yields an average 55% improvement in adversarial accuracy. Moreover, DeepProv complements existing defenses, achieving substantial gains in adversarial robustness. Beyond robustness, we demonstrate the broader potential of DeepProv as an adaptable system to characterize DNN behavior in other critical areas, such as privacy auditing and fairness analysis.

📄 PDF Abstract BibTeX arXiv:2509.26562

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial Robustness

Similar Papers 제목 키워드 기반

VISTA: Knowledge-Driven Vessel Trajectory Imputation with Repair Provenance

2026-01-11 · Hengyu Liu, Tianyi Li, Haoyu Wang, Kristian Torp 외 arxiv

Repairing incomplete trajectory data is essential for downstream spatio-temporal applications. Yet, existing repair methods focus solely on reconstruction without documenting the reasoning behind repair decisions, underm…

Anomaly Detection

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

2026-07-30 · Enjun Du, Hange Zhou, Chenxu Du, Siyi Liu 외 arxiv

Multimodal agents for visual question answering increasingly operate as multi-step trajectories that interleave perception, retrieval, and reasoning, yet evaluation still largely reduces to final-answer accuracy. This ag…

Visual Question AnsweringMultimodal Reasoning

Choices and their Provenance: Explaining Stable Solutions of Abstract Argumentation Frameworks

2025-06-01 · Bertram Ludäscher, Yilin Xia, Shawn Bowers

The rule $\mathrm{Defeated}(x) \leftarrow \mathrm{Attacks}(y,x),\, \neg \, \mathrm{Defeated}(y)$, evaluated under the well-founded semantics (WFS), yields a unique 3-valued (skeptical) solution of an abstract argumentati…

Abstract Argumentation

Reasoning Provenance for Autonomous AI Agents: Structured Behavioral Analytics Beyond State Checkpoints and Execution Traces

2026-03-23 · Neelmani Vispute, Aditya Kadam arxiv

As AI agents transition from human-supervised copilots to autonomous platform infrastructure, the ability to analyze their reasoning behavior across populations of investigations becomes a pressing infrastructure require…

ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents

2026-06-16 · Ander Alvarez, Santhiya Rajan, Samuel Mugel, Román Orús arxiv

Tool-using LLM agents increasingly use the Model Context Protocol (MCP) to answer from heterogeneous evidence sources, including search, APIs, databases, clinical records, and formulary tools. Standard factuality metrics…