paper-with-me

홈 › Papers

Efficient Learning With Sine-Activated Low-rank Matrices

2024-03-28 · Yiping Ji, Hemanth Saratchandran, Cameron Gordon, Zeyu Zhang, Simon Lucey

Low-rank decomposition has emerged as a vital tool for enhancing parameter efficiency in neural network architectures, gaining traction across diverse applications in machine learning. These techniques significantly lower the number of parameters, striking a balance between compactness and performance. However, a common challenge has been the compromise between parameter efficiency and the accuracy of the model, where reduced parameters often lead to diminished accuracy compared to their full-rank counterparts. In this work, we propose a novel theoretical framework that integrates a sinusoidal function within the low-rank decomposition process. This approach not only preserves the benefits of the parameter efficiency characteristic of low-rank methods but also increases the decomposition's rank, thereby enhancing model performance. Our method proves to be a plug in enhancement for existing low-rank models, as evidenced by its successful application in Vision Transformers (ViT), Large Language Models (LLMs), Neural Radiance Fields (NeRF) and 3D shape modelling.

📄 PDF Abstract BibTeX arXiv:2403.19243

Code (0)

등록된 구현이 없습니다.

Tasks

3D Shape ModelingNeRF

Similar Papers 제목 키워드 기반

Training invariances and the low-rank phenomenon: beyond linear networks

2022-01-28 · ICLR 2022 4 · Thien Le, Stefanie Jegelka

The implicit bias induced by the training of neural networks has become a topic of rigorous study. In the limit of gradient flow and gradient descent with appropriate step size, it has been shown that when one trains a d…

Evaluating Singular Value Thresholds for DNN Weight Matrices based on Random Matrix Theory

2025-12-15 · Kohei Nishikawa, Koki Shimizu, Hiroki Hashiguchi arxiv

This study evaluates thresholds for removing singular values from singular value decomposition-based low-rank approximations of deep neural network weight matrices. Each weight matrix is modeled as the sum of signal and …

Implicit Spatiotemporal Bandwidth Enhancement Filter by Sine-activated Deep Learning Model for Fast 3D Photoacoustic Tomography

2025-07-28 · I Gede Eka Sulistyawan, Takuro Ishii, Riku Suzuki, Yoshifumi Saijo arxiv

3D photoacoustic tomography (3D-PAT) using high-frequency hemispherical transducers offers near-omnidirectional reception and enhanced sensitivity to the finer structural details encoded in the high-frequency components …

Little By Little: Continual Learning via Self-Activated Sparse Mixture-of-Rank Adaptive Learning

2025-06-26 · Haodong Lu, Chongyang Zhao, Jason Xue, Lina Yao 외

Continual learning (CL) with large pre-trained models is challenged by catastrophic forgetting and task interference. Existing LoRA-based Mixture-of-Experts (MoE) approaches mitigate forgetting by assigning and freezing …

Continual LearningMixture-of-Experts

Computational ghost imaging with hybrid transforms by integrating Hadamard, discrete cosine, and Haar matrices

2024-05-06 · Yi-Ning Zhao, Lin-Shan Chen, Liu-Ya Chen, Lingxin Kong 외

A scenario of ghost imaging with hybrid transform approach is proposed by integrating Hadamard, discrete cosine, and Haar matrices. The measurement matrix is formed by the Kronecker product of the two different transform…

Image Compression