paper-with-me

Papers

PromptST: Prompt-Enhanced Spatio-Temporal Multi-Attribute Prediction

2023-09-18 · Zijian Zhang, Xiangyu Zhao, Qidong Liu, Chunxu Zhang, Qian Ma, Wanyu Wang, Hongwei Zhao, Yiqi Wang, Zitao Liu

In the era of information explosion, spatio-temporal data mining serves as a critical part of urban management. Considering the various fields demanding attention, e.g., traffic state, human activity, and social event, predicting multiple spatio-temporal attributes simultaneously can alleviate regulatory pressure and foster smart city construction. However, current research can not handle the spatio-temporal multi-attribute prediction well due to the complex relationships between diverse attributes. The key challenge lies in how to address the common spatio-temporal patterns while tackling their distinctions. In this paper, we propose an effective solution for spatio-temporal multi-attribute prediction, PromptST. We devise a spatio-temporal transformer and a parameter-sharing training scheme to address the common knowledge among different spatio-temporal attributes. Then, we elaborate a spatio-temporal prompt tuning strategy to fit the specific attributes in a lightweight manner. Through the pretrain and prompt tuning phases, our PromptST is able to enhance the specific spatio-temoral characteristic capture by prompting the backbone model to fit the specific target attribute while maintaining the learned common knowledge. Extensive experiments on real-world datasets verify that our PromptST attains state-of-the-art performance. Furthermore, we also prove PromptST owns good transferability on unseen spatio-temporal attributes, which brings promising application potential in urban computing. The implementation code is available to ease reproducibility.

📄 PDF Abstract BibTeX arXiv:2309.09500

Code (0)

등록된 구현이 없습니다.

Tasks

AttributePrediction

Similar Papers 제목 키워드 기반

PromptStyler: Prompt-driven Style Generation for Source-free Domain Generalization

2023-07-27 · ICCV 2023 1 · Junhyeong Cho, Gilhyun Nam, Sungyeon Kim, Hunmin Yang 외

In a joint vision-language space, a text feature (e.g., from "a photo of a dog") could effectively represent its relevant image features (e.g., from dog photos). Also, a recent study has demonstrated the cross-modal tran…

Domain GeneralizationImage ClassificationMulti-modal ClassificationMultimodal Deep Learning+4

Prompt Stealing Attacks Against Text-to-Image Generation Models

2023-02-20 · Xinyue Shen, Yiting Qu, Michael Backes, Yang Zhang

Text-to-Image generation models have revolutionized the artwork design process and enabled anyone to create high-quality images by entering text descriptions called prompts. Creating a high-quality prompt that consists o…

Image GenerationText to Image GenerationText-to-Image Generation

Spatio-Temporal Data Enhanced Vision-Language Model for Traffic Scene Understanding

2025-11-12 · Jingtian Ma, Jingyuan Wang, Wayne Xin Zhao, Guoping Liu 외 arxiv

Nowadays, navigation and ride-sharing apps have collected numerous images with spatio-temporal data. A core technology for analyzing such images, associated with spatiotemporal information, is Traffic Scene Understanding…

Scene UnderstandingFew-Shot Learning

PromptStereo: Zero-Shot Stereo Matching via Structure and Motion Prompts

2026-03-02 · Xianqi Wang, Hao Yang, Hangtian Wang, Junda Cheng 외 arxiv

Modern stereo matching methods have leveraged monocular depth foundation models to achieve superior zero-shot generalization performance. However, most existing methods primarily focus on extracting robust features for c…

Zero-shot Generalization

DPStyler: Dynamic PromptStyler for Source-Free Domain Generalization

2024-03-25 · Yunlong Tang, Yuxuan Wan, Lei Qi, Xin Geng

Source-Free Domain Generalization (SFDG) aims to develop a model that works for unseen target domains without relying on any source domain. Research in SFDG primarily bulids upon the existing knowledge of large-scale vis…

Domain GeneralizationSource-free Domain GeneralizationStyle Transfer