paper-with-me

홈 › Papers

Unlocking Temporal Question Answering for Large Language Models with Tailor-Made Reasoning Logic

2023-05-24 · Xingxuan Li, Liying Cheng, Qingyu Tan, Hwee Tou Ng, Shafiq Joty, Lidong Bing

The temporal aspect is a significant dimension of our reality. We notice the challenge that large language models (LLMs) face when engaging in temporal reasoning. Our preliminary experiments show that methods involving the generation of intermediate reasoning steps, such as chain-of-thought and program-aided language models, do not consistently boost the performance of complex temporal question-answering tasks. This limitation can be attributed to the LLMs' inadequate understanding of temporal information. To address this problem, we propose TempLogic, a novel framework designed specifically for temporal question-answering tasks across three levels of reasoning. TempLogic incorporates retrieval-guided context distillation, temporal data extraction, and tailor-made logic reasoning. Extensive experiments and analysis demonstrate the effectiveness of our framework in solving intricate time-bound reasoning tasks.

📄 PDF Abstract BibTeX arXiv:2305.15014

Code (1)

damo-nlp-sg/mvcr 공식 구현 pytorch

Tasks

Logical ReasoningMathQuestion AnsweringRetrieval

Similar Papers 제목 키워드 기반

Unlocking Video-LLM via Agent-of-Thoughts Distillation

2024-12-02 · Yudi Shi, Shangzhe Di, Qirui Chen, Weidi Xie

This paper tackles the problem of video question answering (VideoQA), a task that often requires multi-step reasoning and a profound understanding of spatial-temporal dynamics. While large video-language models perform w…

Language ModelingLanguage ModellingLarge Language ModelMultiple-choice+2

Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks

2024-02-13 · Jusung Lee, Sungguk Cha, Younghyun Lee, Cheoljong Yang

Having revolutionized natural language processing (NLP) applications, large language models (LLMs) are expanding into the realm of multimodal inputs. Owing to their ability to interpret images, multimodal LLMs (MLLMs) ha…

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Model+3

Unlocking UML Class Diagram Understanding in Vision Language Models

2026-05-12 · Artem Naboichenko, René Peinl arxiv

Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard-ing diagrams compared to photos. Although progress has been made in…

Visual Question Answering

ComplexTempQA: A Large-Scale Dataset for Complex Temporal Question Answering

2024-06-07 · Raphael Gruber, Abdelrahman Abdallah, Michael Färber, Adam Jatowt

We introduce ComplexTempQA, a large-scale dataset consisting of over 100 million question-answer pairs designed to tackle the challenges in temporal question answering. ComplexTempQA significantly surpasses existing benc…

Information RetrievalQuestion Answering

LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding

2025-01-09 · Jiaxing Zhao, Boyuan Sun, Xiang Chen, Xihan Wei 외

In this paper, we introduce LLaVA-Octopus, a novel video multimodal large language model. LLaVA-Octopus adaptively weights features from different visual projectors based on user instructions, enabling us to leverage the…

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Model+3