paper-with-me

홈 › Papers

RotRNN: Modelling Long Sequences with Rotations

2024-07-09 · Kai Biegun, Rares Dolga, Jake Cunningham, David Barber

Linear recurrent neural networks, such as State Space Models (SSMs) and Linear Recurrent Units (LRUs), have recently shown state-of-the-art performance on long sequence modelling benchmarks. Despite their success, their empirical performance is not well understood and they come with a number of drawbacks, most notably their complex initialisation and normalisation schemes. In this work, we address some of these issues by proposing RotRNN -- a linear recurrent model which utilises the convenient properties of rotation matrices. We show that RotRNN provides a simple and efficient model with a robust normalisation procedure, and a practical implementation that remains faithful to its theoretical derivation. RotRNN also achieves competitive performance to state-of-the-art linear recurrent models on several long sequence modelling datasets.

📄 PDF Abstract BibTeX arXiv:2407.07239

Code (0)

등록된 구현이 없습니다.

Tasks

State Space Models

Similar Papers 제목 키워드 기반

Hamiltonian latent operators for content and motion disentanglement in image sequences

2021-12-02 · Asif Khan, Amos Storkey

We introduce \textit{HALO} -- a deep generative model utilising HAmiltonian Latent Operators to reliably disentangle content and motion information in image sequences. The \textit{content} represents summary statistics o…

DisentanglementMotion Disentanglement

On the Generation of Long Binary Sequences with Record-Breaking PSL Values

2021-04-02 · Miroslav Dimitrov, Tsonka Baitcheva, Nikolay Nikolov

Binary sequences are widely used in various practical fields, such as telecommunications, radar technology, navigation, cryptography, measurement sciences, biology or industry. In this paper, a method to generate long bi…

HoPE: Hyperbolic Rotary Positional Encoding for Stable Long-Range Dependency Modeling in Large Language Models

2025-09-05 · Chang Dai, Hongyu Shan, Mingyang Song, Di Liang arxiv

Positional encoding mechanisms enable Transformers to model sequential structure and long-range dependencies in text. While absolute positional encodings struggle with extrapolation to longer sequences due to fixed posit…

MotioNet: 3D Human Motion Reconstruction from Monocular Video with Skeleton Consistency

2020-06-22 · Mingyi Shi, Kfir Aberman, Andreas Aristidou, Taku Komura 외

We introduce MotioNet, a deep neural network that directly reconstructs the motion of a 3D human skeleton from monocular video.While previous methods rely on either rigging or inverse kinematics (IK) to associate a consi…

QuaterNet: A Quaternion-based Recurrent Model for Human Motion

2018-05-16 · Dario Pavllo, David Grangier, Michael Auli

Deep learning for predicting or generating 3D human pose sequences is an active research area. Previous work regresses either joint rotations or joint positions. The former strategy is prone to error accumulation along t…

3D Human Pose EstimationMotion EstimationPosition