paper-with-me

홈 › Papers

TempoFormer: A Transformer for Temporally-aware Representations in Change Detection

2024-08-28 · Talia Tseriotou, Adam Tsakalidis, Maria Liakata

Dynamic representation learning plays a pivotal role in understanding the evolution of linguistic content over time. On this front both context and time dynamics as well as their interplay are of prime importance. Current approaches model context via pre-trained representations, which are typically temporally agnostic. Previous work on modelling context and temporal dynamics has used recurrent methods, which are slow and prone to overfitting. Here we introduce TempoFormer, the first task-agnostic transformer-based and temporally-aware model for dynamic representation learning. Our approach is jointly trained on inter and intra context dynamics and introduces a novel temporal variation of rotary positional embeddings. The architecture is flexible and can be used as the temporal representation foundation of other models or applied to different transformer-based architectures. We show new SOTA performance on three different real-time change detection tasks.

📄 PDF Abstract BibTeX arXiv:2408.15689

Code (0)

등록된 구현이 없습니다.

Tasks

Change DetectionRepresentation Learning

Similar Papers 제목 키워드 기반

Audio Visual Scene-Aware Dialog Generation with Transformer-based Video Representations

2022-02-21 · Yoshihiro Yamazaki, Shota Orihashi, Ryo Masumura, Mihiro Uchida 외

There have been many attempts to build multimodal dialog systems that can respond to a question about given audio-visual information, and the representative task for such systems is the Audio Visual Scene-Aware Dialog (A…

Answer GenerationVideo Understanding

Temporal Attention for Language Models

2022-02-04 · Findings (NAACL) 2022 7 · Guy D. Rosin, Kira Radinsky

Pretrained language models based on the transformer architecture have shown great success in NLP. Textual training data often comes from the web and is thus tagged with time-specific information, but most language models…

Change Detection

VLA-4D: Embedding 4D Awareness into Vision-Language-Action Models for SpatioTemporally Coherent Robotic Manipulation

2025-11-21 · Hanyu Zhou, Chuanhao Ma, Gim Hee Lee arxiv

Vision-language-action (VLA) models show potential for general robotic tasks, but remain challenging in spatiotemporally coherent manipulation, which requires fine-grained representations. Typically, existing methods emb…

Uncertainty-Aware Token Importance Estimation in Spiking Transformers

2026-05-10 · Wenxuan Liu, Zecheng Hao, Tong Bu, Yuran Wang 외 arxiv

Spiking transformers have shown strong potential for neuromorphic vision, yet their token processing across multiple spiking steps still introduces substantial redundancy and inference cost. Existing token reduction meth…

Words with Consistent Diachronic Usage Patterns are Learned Earlier: A Computational Analysis Using Temporally Aligned Word Embeddings

2021-04-20 · Cognitive Science 2021 4 · Giovanni Cassani, Federico Bianchi, Marco Marelli

In this study, we use temporally aligned word embeddings and a large diachronic corpus of English to quantify language change in a data-driven, scalable way, which is grounded in language use. We show a unique and reliab…

Diachronic Word EmbeddingsDiversityRelationWord Embeddings