paper-with-me

홈 › Papers

GiVA: Gradient-Informed Bases for Vector-Based Adaptation

2026-04-23 · Neeraj Gangwar, Rishabh Deshmukh, Michael Shavlovsky, Hancao Li, Vivek Mittal, Lexing Ying, Nickvash Kani arxiv

As model sizes continue to grow, parameter-efficient fine-tuning has emerged as a powerful alternative to full fine-tuning. While LoRA is widely adopted among these methods, recent research has explored vector-based adaptation methods due to their extreme parameter efficiency. However, these methods typically require substantially higher ranks than LoRA to match its performance, leading to increased training costs. This work introduces GiVA, a gradient-based initialization strategy for vector-based adaptation. It achieves training times comparable to LoRA and maintains the extreme parameter efficiency of vector-based adaptation. We evaluate GiVA across diverse benchmarks, including natural language understanding, natural language generation, and image classification. Experiments show that our approach consistently outperforms or achieves performance competitive with existing vector-based adaptation methods and LoRA while reducing rank requirements by a factor of eight ($8\times$).

📄 PDF Abstract BibTeX arXiv:2604.21901

Code (0)

등록된 구현이 없습니다.

Tasks

parameter-efficient fine-tuningNatural Language UnderstandingImage Classification

Similar Papers 제목 키워드 기반

Unified Policy Value Decomposition for Rapid Adaptation

2026-03-18 · Cristiano Capone, Luca Falorsi, Andrea Ciardiello, Luca Manneschi arxiv

Rapid adaptation in complex control systems remains a central challenge in reinforcement learning. We introduce a framework in which policy and value functions share a low-dimensional coefficient vector - a goal embeddin…

Reinforcement Learning

Basis-Oriented Low-rank Transfer for Few-Shot and Test-Time Adaptation

2025-12-02 · Junghwan Park, Woojin Cho, Junhyuk Heo, Darongsae Kwon 외 arxiv

Adapting large pre-trained models to unseen tasks under tight data and compute budgets remains challenging. Meta-learning approaches explicitly learn good initializations, but they require an additional meta-training pha…

parameter-efficient fine-tuningTest-time Adaptation

TFormer: 3D Tooth Segmentation in Mesh Scans with Geometry Guided Transformer

2022-10-29 · Huimin Xiong, Kunle Li, Kaiyuan Tan, Yang Feng 외

Optical Intra-oral Scanners (IOS) are widely used in digital dentistry, providing 3-Dimensional (3D) and high-resolution geometrical information of dental crowns and the gingiva. Accurate 3D tooth segmentation, which aim…

Multi-Task LearningSegmentation

Dimension reduction for derivative-informed operator learning: An analysis of approximation errors

2025-04-11 · Dingcheng Luo, Thomas O'Leary-Roseberry, Peng Chen, Omar Ghattas

We study the derivative-informed learning of nonlinear operators between infinite-dimensional separable Hilbert spaces by neural networks. Such operators can arise from the solution of partial differential equations (PDE…

Dimensionality ReductionExperimental DesignOperator learning

Locally adaptive activation functions with slope recovery term for deep and physics-informed neural networks

2019-09-25 · Ameya D. Jagtap, Kenji Kawaguchi, George Em. Karniadakis

We propose two approaches of locally adaptive activation functions namely, layer-wise and neuron-wise locally adaptive activation functions, which improve the performance of deep and physics-informed neural networks. The…

Data Augmentation