paper-with-me

홈 › Papers

Neural Sequence-to-grid Module for Learning Symbolic Rules

2021-01-13 · Segwang Kim, Hyoungwook Nam, Joonyoung Kim, Kyomin Jung

Logical reasoning tasks over symbols, such as learning arithmetic operations and computer program evaluations, have become challenges to deep learning. In particular, even state-of-the-art neural networks fail to achieve \textit{out-of-distribution} (OOD) generalization of symbolic reasoning tasks, whereas humans can easily extend learned symbolic rules. To resolve this difficulty, we propose a neural sequence-to-grid (seq2grid) module, an input preprocessor that automatically segments and aligns an input sequence into a grid. As our module outputs a grid via a novel differentiable mapping, any neural network structure taking a grid input, such as ResNet or TextCNN, can be jointly trained with our module in an end-to-end fashion. Extensive experiments show that neural networks having our module as an input preprocessor achieve OOD generalization on various arithmetic and algorithmic problems including number sequence prediction problems, algebraic word problems, and computer program evaluation problems while other state-of-the-art sequence transduction models cannot. Moreover, we verify that our module enhances TextCNN to solve the bAbI QA tasks without external memory.

📄 PDF Abstract BibTeX arXiv:2101.04921

Code (1)

SegwangKim/neural-seq2grid-module 공식 구현 tf

Tasks

Logical Reasoning

Methods 이 논문이 사용한 방법론

Average Pooling 설명 없음
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Residual Connection 설명 없음
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Batch Normalization 설명 없음
Residual Block Residual Blocks are skip-connection blocks that learn residual functions with reference to the layer inputs, instead of learning unreferenced functions. They were introduced…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

Improved Logical Reasoning of Language Models via Differentiable Symbolic Programming

2023-05-05 · HANLIN ZHANG, Jiani Huang, Ziyang Li, Mayur Naik 외

Pre-trained large language models (LMs) struggle to perform logical reasoning reliably despite advances in scale and compositionality. In this work, we tackle this challenge through the lens of symbolic programming. We p…

Logical Reasoning

Slots, Transitions, Loops: Learning Composable World Models for ARC

2026-06-10 · Gege Gao, Bernhard Schölkopf, Andreas Geiger arxiv

ARC tests in-context rule induction: given a few input-output demonstrations, a model must infer the hidden rule and apply it to a new query. While many approaches express ARC rules through language, code, or symbolic pr…

Learning Symbolic Rules for Interpretable Deep Reinforcement Learning

2021-03-15 · Zhihao Ma, Yuzheng Zhuang, Paul Weng, Hankz Hankui Zhuo 외

Recent progress in deep reinforcement learning (DRL) can be largely attributed to the use of neural networks. However, this black-box approach fails to explain the learned policy in a human understandable way. To address…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Perform Like an Engine: A Closed-Loop Neural-Symbolic Learning Framework for Knowledge Graph Inference

2021-12-02 · COLING 2022 10 · Guanglin Niu, Bo Li, Yongfei Zhang, ShiLiang Pu

Knowledge graph (KG) inference aims to address the natural incompleteness of KGs, including rule learning-based and KG embedding (KGE) models. However, the rule learning-based models suffer from low efficiency and genera…

Link Prediction

Regression Planning Networks

2019-09-28 · NeurIPS 2019 12 · Danfei Xu, Roberto Martín-Martín, De-An Huang, Yuke Zhu 외

Recent learning-to-plan methods have shown promising results on planning directly from observation space. Yet, their ability to plan for long-horizon tasks is limited by the accuracy of the prediction model. On the other…

regression