paper-with-me

홈 › Papers

LoTR: Low Tensor Rank Weight Adaptation

2024-02-02 · Daniel Bershatsky, Daria Cherniuk, Talgat Daulbaev, Aleksandr Mikhalev, Ivan Oseledets

In this paper we generalize and extend an idea of low-rank adaptation (LoRA) of large language models (LLMs) based on Transformer architecture. Widely used LoRA-like methods of fine-tuning LLMs are based on matrix factorization of gradient update. We introduce LoTR, a novel approach for parameter-efficient fine-tuning of LLMs which represents a gradient update to parameters in a form of tensor decomposition. Low-rank adapter for each layer is constructed as a product of three matrices, and tensor structure arises from sharing left and right multipliers of this product among layers. Simultaneous compression of a sequence of layers with low-rank tensor representation allows LoTR to archive even better parameter efficiency then LoRA especially for deep models. Moreover, the core tensor does not depend on original weight dimension and can be made arbitrary small, which allows for extremely cheap and fast downstream fine-tuning.

📄 PDF Abstract BibTeX arXiv:2402.01376

Code (0)

등록된 구현이 없습니다.

Tasks

parameter-efficient fine-tuningTensor Decomposition

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Multi-Head Attention 설명 없음
Residual Connection 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…

Similar Papers 제목 키워드 기반

Low Tensor-Rank Adaptation of Kolmogorov--Arnold Networks

2025-02-10 · Yihang Gao, Michael K. Ng, Vincent Y. F. Tan

Kolmogorov--Arnold networks (KANs) have demonstrated their potential as an alternative to multi-layer perceptions (MLPs) in various domains, especially for science-related tasks. However, transfer learning of KANs remain…

image-classificationImage ClassificationKolmogorov-Arnold NetworksTransfer Learning

LORA-CRAFT: Cross-layer Rank Adaptation via Frozen Tucker Decomposition of Pre-trained Attention Weights

2026-02-19 · Kasun Dewage, Marianna Pensky, Suranadi De Silva, Shankadeep Mondal arxiv

We introduce CRAFT (Cross-layer Rank Adaptation via Frozen Tucker), a parameter-efficient fine-tuning (PEFT) method that applies Tucker tensor decomposition to pre-trained attention weight matrices stacked across transfo…

parameter-efficient fine-tuning

DoTA: Weight-Decomposed Tensor Adaptation for Large Language Models

2024-12-30 · Xiaolin Hu, Xiang Cheng, Peiyu Liu, Wei Liu 외

Low-rank adaptation (LoRA) reduces the computational and memory demands of fine-tuning large language models (LLMs) by approximating updates with low-rank matrices. However, low-rank approximation in two-dimensional spac…

Arithmetic ReasoningQuantizationTensor Decomposition

TeRA: Vector-based Random Tensor Network for High-Rank Adaptation of Large Language Models

2025-09-03 · Yuxuan Gu, Wuyang Zhou, Giorgos Iacovides, Danilo Mandic arxiv

Parameter-Efficient Fine-Tuning (PEFT) methods, such as Low-Rank Adaptation (LoRA), have significantly reduced the number of trainable parameters needed in fine-tuning large language models (LLMs). The developments of Lo…

parameter-efficient fine-tuning

Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image Models

2025-01-15 · Zerui Tao, Yuhta Takida, Naoki Murata, Qibin Zhao 외

Parameter-Efficient Fine-Tuning (PEFT) of text-to-image models has become an increasingly popular technique with many applications. Among the various PEFT methods, Low-Rank Adaptation (LoRA) and its variants have gained …

parameter-efficient fine-tuningTensor Decomposition