paper-with-me

홈 › Papers

BiTimeBERT: Extending Pre-Trained Language Representations with Bi-Temporal Information

2022-04-27 · Jiexin Wang, Adam Jatowt, Masatoshi Yoshikawa, Yi Cai

Time is an important aspect of documents and is used in a range of NLP and IR tasks. In this work, we investigate methods for incorporating temporal information during pre-training to further improve the performance on time-related tasks. Compared with common pre-trained language models like BERT which utilize synchronic document collections (e.g., BookCorpus and Wikipedia) as the training corpora, we use long-span temporal news article collection for building word representations. We introduce BiTimeBERT, a novel language representation model trained on a temporal collection of news articles via two new pre-training tasks, which harnesses two distinct temporal signals to construct time-aware language representations. The experimental results show that BiTimeBERT consistently outperforms BERT and other existing pre-trained models with substantial gains on different downstream NLP tasks and applications for which time is of importance (e.g., the accuracy improvement over BERT is 155\% on the event time estimation task).

📄 PDF Abstract BibTeX arXiv:2204.13032

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesInformation Retrieval

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Multi-Head Attention 설명 없음
Residual Connection 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
WordPiece 설명 없음
Weight Decay 설명 없음
Adam 설명 없음

Similar Papers 제목 키워드 기반

Towards Effective Time-Aware Language Representation: Exploring Enhanced Temporal Understanding in Language Models

2024-06-04 · Jiexin Wang, Adam Jatowt, Yi Cai

In the evolving field of Natural Language Processing, understanding the temporal context of text is increasingly crucial. This study investigates methods to incorporate temporal information during pre-training, aiming to…

Document DatingLanguage ModelingLanguage ModellingMasked Language Modeling

When Machines Speak: A Unified Generative Framework for Integrating Machine-Native Symbols into Pretrained Large Language Models

2026-08-20 · Su Yan, Rakesh Iyer arxiv

Many real-world AI systems represent entities, behaviors, and structured information using discrete machine-native symbols rather than natural language. While these representations are compact and preserve task-relevant …

Sequential RecommendationStructured Prediction

Learning Parallel Dense Correspondence from Spatio-Temporal Descriptors for Efficient and Robust 4D Reconstruction

2021-03-30 · CVPR 2021 1 · Jiapeng Tang, Dan Xu, Kui Jia, Lei Zhang

This paper focuses on the task of 4D shape reconstruction from a sequence of point clouds. Despite the recent success achieved by extending deep implicit representations into 4D space, it is still a great challenge in tw…

4D reconstruction

CITRIS: Causal Identifiability from Temporal Intervened Sequences

2022-02-07 · Phillip Lippe, Sara Magliacane, Sindy Löwe, Yuki M. Asano 외

Understanding the latent causal factors of a dynamical system from visual observations is considered a crucial step towards agents reasoning in complex environments. In this paper, we propose CITRIS, a variational autoen…

Representation LearningTemporal Sequences

Espresso: High Compression For Rich Extraction From Videos for Your Vision-Language Model

2024-12-06 · Keunwoo Peter Yu, Achal Dave, Rares Ambrus, Jean Mercat

Recent advances in vision-language models (VLMs) have shown great promise in connecting images and text, but extending these models to long videos remains challenging due to the rapid growth in token counts. Models that …

EgoSchemaLanguage ModelingLanguage ModellingVideo Understanding