Addressing the Scarcity of Benchmarks for Graph XAI
While Graph Neural Networks (GNNs) have become the de facto model for learning from structured data, their decisional process remains opaque to the end user, undermining their deployment in safety-critical applications. In the case of graph classification, Explainable Artificial Intelligence (XAI) techniques address this major issue by identifying sub-graph motifs that explain predictions. However, advancements in this field are hindered by a chronic scarcity of benchmark datasets with known ground-truth motifs to assess the explanations' quality. Current graph XAI benchmarks are limited to synthetic data or a handful of real-world tasks hand-curated by domain experts. In this paper, we propose a general method to automate the construction of XAI benchmarks for graph classification from real-world datasets. We provide both 15 ready-made benchmarks, as well as the code to generate more than 2000 additional XAI benchmarks with our method. As a use case, we employ our benchmarks to assess the effectiveness of some popular graph explainers.
Code (1)
Tasks
Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Graph ClassificationSimilar Papers 제목 키워드 기반
Homophily Enhanced Graph Domain Adaptation
Graph Domain Adaptation (GDA) transfers knowledge from labeled source graphs to unlabeled target graphs, addressing the challenge of label scarcity. In this paper, we highlight the significance of graph homophily, a pivo…
Domain AdaptationGRAPH DOMAIN ADAPTATIONWhen Do Contrastive Learning Signals Help Spatio-Temporal Graph Forecasting?
Deep learning models are modern tools for spatio-temporal graph (STG) forecasting. Though successful, we argue that data scarcity is a key factor limiting their recent improvements. Meanwhile, contrastive learning has be…
Contrastive LearningData AugmentationSemantic SimilaritySemantic Textual SimilarityCalliReader: Contextualizing Chinese Calligraphy via an Embedding-Aligned Vision-Language Model
Chinese calligraphy, a UNESCO Heritage, remains computationally challenging due to visual ambiguity and cultural complexity. Existing AI systems fail to contextualize their intricate scripts, because of limited annotated…
HallucinationLanguage ModelingLanguage ModellingOptical Character Recognition (OCR)+1Synthetic Datasets for Machine Learning on Spatio-Temporal Graphs using PDEs
Many physical processes can be expressed through partial differential equations (PDEs). Real-world measurements of such processes are often collected at irregularly distributed points in space, which can be effectively r…
BenchmarkingEpidemiologyData Scarcity in Recommendation Systems: A Survey
The prevalence of online content has led to the widespread adoption of recommendation systems (RSs), which serve diverse purposes such as news, advertisements, and e-commerce recommendations. Despite their significance, …
Data AugmentationRecommendation SystemsSelf-Supervised LearningSurvey+1