paper-with-me

홈 › Papers

Joint Spatio-Textual Reasoning for Answering Tourism Questions

2020-09-28 · Danish Contractor, Shashank Goel, Mausam, Parag Singla

Our goal is to answer real-world tourism questions that seek Points-of-Interest (POI) recommendations. Such questions express various kinds of spatial and non-spatial constraints, necessitating a combination of textual and spatial reasoning. In response, we develop the first joint spatio-textual reasoning model, which combines geo-spatial knowledge with information in textual corpora to answer questions. We first develop a modular spatial-reasoning network that uses geo-coordinates of location names mentioned in a question, and of candidate answer POIs, to reason over only spatial constraints. We then combine our spatial-reasoner with a textual reasoner in a joint model and present experiments on a real world POI recommendation task. We report substantial improvements over existing models with-out joint spatio-textual reasoning.

📄 PDF Abstract BibTeX arXiv:2009.13613

Code (1)

dair-iitd/TourismQA 공식 구현

Tasks

Spatial Reasoning

Similar Papers 제목 키워드 기반

Location Aware Modular Biencoder for Tourism Question Answering

2024-01-04 · Haonan Li, Martin Tomko, Timothy Baldwin

Answering real-world tourism questions that seek Point-of-Interest (POI) recommendations is challenging, as it requires both spatial and non-spatial reasoning, over a large candidate pool. The traditional method of encod…

Question AnsweringRetrievalSpatial Reasoning

A novel forecasting framework combining virtual samples and enhanced Transformer models for tourism demand forecasting

2025-03-25 · Tingting Diao, Xinzhang Wu, Lina Yang, Ling Xiao 외

Accurate tourism demand forecasting is hindered by limited historical data and complex spatiotemporal dependencies among tourist origins. A novel forecasting framework integrating virtual sample generation and a novel Tr…

Demand ForecastingManagement

Track the Answer: Extending TextVQA from Image to Video with Spatio-Temporal Clues

2024-12-17 · Yan Zhang, Gangyan Zeng, Huawen Shen, Daiqing Wu 외

Video text-based visual question answering (Video TextVQA) is a practical task that aims to answer questions by jointly reasoning textual and visual information in a given video. Inspired by the development of TextVQA in…

Language ModelingLanguage ModellingOptical Character Recognition (OCR)Question Answering+2

VL-MemKnG: Hybrid Memory with a Spatio-Temporal Knowledge Graph for Question Answering over Long Egocentric Navigation Trajectories

2026-06-15 · Svetlana Lukina, Mohamad Al Mdfaa, Gloria Haro, Sergey Zagoruyko 외 arxiv

Answering navigation-relevant questions over long egocentric videos requires retrieving and organizing evidence distributed across distant temporal moments while maintaining spatial and contextual consistency. Although l…

Video Question AnsweringKnowledge Graphs

STRIVE: Structured Spatiotemporal Exploration for Reinforcement Learning in Video Question Answering

2026-04-02 · Emad Bahrami, Olga Zatsarynna, Parth Pathak, Sunando Sengupta 외 arxiv

We introduce STRIVE (SpatioTemporal Reinforcement with Importance-aware Variant Exploration), a structured reinforcement learning framework for video question answering. While group-based policy optimization methods have…

Video Question AnsweringReinforcement Learning