paper-with-me

홈 › Papers

Harnessing Rich Multi-Modal Data for Spatial-Temporal Homophily-Embedded Graph Learning Across Domains and Localities

2025-12-11 · Takuya Kurihana, Xiaojian Zhang, Wing Yee Au, Hon Yung Wong arxiv

Modern cities are increasingly reliant on data-driven insights to support decision making in areas such as transportation, public safety and environmental impact. However, city-level data often exists in heterogeneous formats, collected independently by local agencies with diverse objectives and standards. Despite their numerous, wide-ranging, and uniformly consumable nature, national-level datasets exhibit significant heterogeneity and multi-modality. This research proposes a heterogeneous data pipeline that performs cross-domain data fusion over time-varying, spatial-varying and spatial-varying time-series datasets. We aim to address complex urban problems across multiple domains and localities by harnessing the rich information over 50 data sources. Specifically, our data-learning module integrates homophily from spatial-varying dataset into graph-learning, embedding information of various localities into models. We demonstrate the generalizability and flexibility of the framework through five real-world observations using a variety of publicly accessible datasets (e.g., ride-share, traffic crash, and crime reports) collected from multiple cities. The results show that our proposed framework demonstrates strong predictive performance while requiring minimal reconfiguration when transferred to new localities or domains. This research advances the goal of building data-informed urban systems in a scalable way, addressing one of the most pressing challenges in smart city analytics.

📄 PDF Abstract BibTeX arXiv:2512.11178

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingGraph Learning

Similar Papers 제목 키워드 기반

FedMMKT:Co-Enhancing a Server Text-to-Image Model and Client Task Models in Multi-Modal Federated Learning

2025-10-14 · Ningxin He, Yang Liu, Wei Sun, Xiaozhou Ye 외 arxiv

Text-to-Image (T2I) models have demonstrated their versatility in a wide range of applications. However, adaptation of T2I models to specialized tasks is often limited by the availability of task-specific data due to pri…

Federated Learning

Harnessing Webpage UIs for Text-Rich Visual Understanding

2024-10-17 · Junpeng Liu, Tianyue Ou, YiFan Song, Yuxiao Qu 외

Text-rich visual understanding-the ability to process environments where dense textual content is integrated with visuals-is crucial for multimodal large language models (MLLMs) to interact effectively with structured en…

document understandingOptical Character Recognition (OCR)

Multi-Modality Spatio-Temporal Forecasting via Self-Supervised Learning

2024-05-06 · Jiewen Deng, Renhe Jiang, JiaQi Zhang, Xuan Song

Multi-modality spatio-temporal (MoST) data extends spatio-temporal (ST) data by incorporating multiple modalities, which is prevalent in monitoring systems, encompassing diverse traffic demands and air quality assessment…

Self-Supervised LearningSpatio-Temporal Forecasting

MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning

2025-07-16 · Hongxu Ma, Guanshuo Wang, Fufu Yu, Qiong Jia 외 arxiv

Video Moment Retrieval (MR) and Highlight Detection (HD) aim to pinpoint specific moments and assess clip-wise relevance based on the text query. While DETR-based joint frameworks have made significant strides, there rem…

Highlight DetectionMoment Retrieval

Harnessing the Potential of Spatial Statistics for Spatial Omics Data with pasta

2024-12-02 · Martin Emons, Samuel Gunz, Helena L. Crowell, Izaskun Mallona 외

Spatial omics assays allow for the molecular characterisation of cells in their spatial context. Notably, the two main technological streams, imaging-based and high-throughput sequencing-based, can give rise to very diff…