paper-with-me

홈 › Papers

Text-Guided Multi-Scale Frequency Representation Adaptation

2026-05-05 · Weicai Yan, Xinhua Ma, Wang Lin, Tao Jin arxiv

Parameter-efficient fine-tuning methods introduce a small number of training parameters, enabling pre-trained models to adapt rapidly to new data distributions. While these methods have shown promising results, they exhibit notable limitations. First, most existing methods operate in the signal space domain, which results in substantial information redundancy. Second, most existing methods utilize fixed prompts or adaptation layers, failing to fully account for the multi-scale characteristics of signals. To address these challenges, we propose the Multi-Scale Frequency Adapter (FreqAdapter), which integrates textual information and performs multi-scale fine-tuning of signals in the frequency domain. Additionally, we introduce a multi-scale adaptation strategy to optimize receptive fields across different frequency ranges, further enhancing the model's representational capacity. Extensive experiments on multimodal models, including CLIP and LLaVA, demonstrate that FreqAdapter significantly improves both performance and efficiency. FreqAdapter improves performance with minimal cost and fast convergence within one epoch. Code is available at https://github.com/Kelvin-ywc/FreqAdapter.

📄 PDF Abstract BibTeX arXiv:2605.08181

Code (0)

등록된 구현이 없습니다.

Tasks

parameter-efficient fine-tuning

Similar Papers 제목 키워드 기반

Speaker Representation Learning using Global Context Guided Channel and Time-Frequency Transformations

2020-09-02 · Wei Xia, John H. L. Hansen

In this study, we propose the global context guided channel and time-frequency transformations to model the long-range, non-local time-frequency dependencies and channel variances in speaker representations. We use the g…

Representation LearningSpeaker Verification

FreqDINO: Frequency-Guided Adaptation for Generalized Boundary-Aware Ultrasound Image Segmentation

2025-12-12 · Yixuan Zhang, Qing Xu, Yue Li, Xiangjian He 외 arxiv

Ultrasound image segmentation is pivotal for clinical diagnosis, yet challenged by speckle noise and imaging artifacts. Recently, DINOv3 has shown remarkable promise in medical image segmentation with its powerful repres…

Medical Image Segmentation

Class-frequency Guided Noise Schedule for Diffusion Models

2026-06-26 · Jiequan Cui, Beier Zhu, Qingshan Xu, Xiaojuan Qi 외 arxiv

In this paper, we are the first to examine the correlations between class frequency and the multi-scale noise schedule within diffusion models. For score-based generative models, low-density regions often lead to inaccur…

Text-to-Image GenerationImage Classification

From Spatial to Spectral: An Efficient, Frequency-Guided Feature Representation Learner for Small Object Detection

2026-06-22 · Yuhan Rui, Shihan Qiao, Yibin Lou, Mingxi Yu 외 arxiv

Efficient small object detection is bottlenecked by the inherent feature scarcity of tiny targets, which is further aggravated by operations of spatial-domain detectors that indiscriminately discard critical high-frequen…

Small Object Detection

Multiscale Structure-Guided Latent Diffusion for Multimodal MRI Translation

2026-03-13 · Jianqiang Lin, Zhiqiang Shen, Peng Cao, Jinzhu Yang 외 arxiv

Although diffusion models have achieved remarkable progress in multi-modal magnetic resonance imaging (MRI) translation tasks, existing methods still tend to suffer from anatomical inconsistencies or degraded texture det…