Evaluating Logical Generalization in Graph Neural Networks
Recent research has highlighted the role of relational inductive biases in building learning agents that can generalize and reason in a compositional manner. However, while relational learning algorithms such as graph neural networks (GNNs) show promise, we do not understand how effectively these approaches can adapt to new tasks. In this work, we study the task of logical generalization using GNNs by designing a benchmark suite grounded in first-order logic. Our benchmark suite, GraphLog, requires that learning algorithms perform rule induction in different synthetic logics, represented as knowledge graphs. GraphLog consists of relation prediction tasks on 57 distinct logical domains. We use GraphLog to evaluate GNNs in three different setups: single-task supervised learning, multi-task pretraining, and continual learning. Unlike previous benchmarks, our approach allows us to precisely control the logical relationship between the different tasks. We find that the ability for models to generalize and adapt is strongly determined by the diversity of the logical rules they encounter during training, and our results highlight new challenges for the design of GNN models. We publicly release the dataset and code used to generate and interact with the dataset at https://www.cs.mcgill.ca/~ksinha4/graphlog.
Code (1)
Tasks
Continual LearningDiversityKnowledge GraphsMulti-Task LearningRelational ReasoningRelation PredictionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Neural Compositional Rule Learning for Knowledge Graph Reasoning
Learning logical rules is critical to improving reasoning in KGs. This is due to their ability to provide logical and interpretable explanations when used for predictions, as well as their ability to generalize to other …
Knowledge Graph CompletionSystematic GeneralizationSize Generalization of Graph Neural Networks on Biological Data: Insights and Practices from the Spectral Perspective
We investigate size-induced distribution shifts in graphs and assess their impact on the ability of graph neural networks (GNNs) to generalize to larger graphs relative to the training data. Existing literature presents …
Graph ClassificationGraphUniverse: Synthetic Graph Generation for Evaluating Inductive Generalization
A fundamental challenge in graph learning is understanding how models generalize to new, unseen graphs. While synthetic benchmarks offer controlled settings for analysis, existing approaches are confined to single-graph,…
Graph GenerationGraph LearningCLUTRR: A Diagnostic Benchmark for Inductive Reasoning from Text
The recent success of natural language understanding (NLU) systems has been troubled by results highlighting the failure of these models to generalize in a systematic and robust way. In this work, we introduce a diagnost…
DiagnosticGraph Neural NetworkInductive logic programmingNatural Language Understanding+2Evaluating GPT-4 with Vision on Detection of Radiological Findings on Chest Radiographs
The study examines the application of GPT-4V, a multi-modal large language model equipped with visual recognition, in detecting radiological findings from a set of 100 chest radiographs and suggests that GPT-4V is curren…
DiagnosticLanguage ModelingLanguage ModellingLarge Language Model