paper-with-me

홈 › Papers

Puzzle Solving using Reasoning of Large Language Models: A Survey

2024-02-17 · Panagiotis Giadikiaroglou, Maria Lymperaiou, Giorgos Filandrianos, Giorgos Stamou

Exploring the capabilities of Large Language Models (LLMs) in puzzle solving unveils critical insights into their potential and challenges in AI, marking a significant step towards understanding their applicability in complex reasoning tasks. This survey leverages a unique taxonomy -- dividing puzzles into rule-based and rule-less categories -- to critically assess LLMs through various methodologies, including prompting techniques, neuro-symbolic approaches, and fine-tuning. Through a critical review of relevant datasets and benchmarks, we assess LLMs' performance, identifying significant challenges in complex puzzle scenarios. Our findings highlight the disparity between LLM capabilities and human-like reasoning, particularly in those requiring advanced logical inference. The survey underscores the necessity for novel strategies and richer datasets to advance LLMs' puzzle-solving proficiency and contribute to AI's logical reasoning and creative problem-solving advancements.

📄 PDF Abstract BibTeX arXiv:2402.11291

Code (0)

등록된 구현이 없습니다.

Tasks

Logical ReasoningSurvey

Similar Papers 제목 키워드 기반

Reasoning or Pattern Matching? Probing Large Vision-Language Models with Visual Puzzles

2026-01-20 · Maria Lymperaiou, Vasileios Karampinis, Giorgos Filandrianos, Angelos Vlachos 외 arxiv

Puzzles have long served as compact and revealing probes of human cognition, isolating abstraction, rule discovery, and systematic reasoning with minimal reliance on prior knowledge. Leveraging these properties, visual p…

Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning

2024-03-06 · Deepanway Ghosal, Vernon Toh Yan Han, Chia Yew Ken, Soujanya Poria

This paper introduces the novel task of multimodal puzzle solving, framed within the context of visual question-answering. We present a new dataset, AlgoPuzzleVQA designed to challenge and evaluate the capabilities of mu…

Multimodal ReasoningQuestion AnsweringVisual Question Answering

VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

2025-03-29 · Yufan Ren, Konstantinos Tertikas, Shalini Maiti, Junlin Han 외

Large Vision-Language Models (LVLMs) struggle with puzzles, which require precise perception, rule comprehension, and logical reasoning. Assessing and enhancing their performance in this domain is crucial, as it reflects…

Logical Reasoning

Logic-of-Thought: Empowering Large Language Models with Logic Programs for Solving Puzzles in Natural Language

2025-05-22 · Naiqi Li, Peiyuan Liu, Zheng Liu, Tao Dai 외

Solving puzzles in natural language poses a long-standing challenge in AI. While large language models (LLMs) have recently shown impressive capabilities in a variety of tasks, they continue to struggle with complex puzz…

Natural Language Understanding

Solving Situation Puzzles with Large Language Model and External Reformulation

2025-03-24 · Kun Li, Xinwei Chen, Tianyou Song, Chengrui Zhou 외

In recent years, large language models (LLMs) have shown an impressive ability to perform arithmetic and symbolic reasoning tasks. However, we found that LLMs (e.g., ChatGPT) cannot perform well on reasoning that require…

Language ModelingLanguage ModellingLarge Language Model