Lost in Translation? Vocabulary Alignment for Source-Free Adaptation in Open-Vocabulary Semantic Segmentation
We introduce VocAlign, a novel source-free domain adaptation framework specifically designed for VLMs in open-vocabulary semantic segmentation. Our method adopts a student-teacher paradigm enhanced with a vocabulary alignment strategy, which improves pseudo-label generation by incorporating additional class concepts. To ensure efficiency, we use Low-Rank Adaptation (LoRA) to fine-tune the model, preserving its original capabilities while minimizing computational overhead. In addition, we propose a Top-K class selection mechanism for the student model, which significantly reduces memory requirements while further improving adaptation performance. Our approach achieves a notable 6.11 mIoU improvement on the CityScapes dataset and demonstrates superior performance on zero-shot segmentation benchmarks, setting a new standard for source-free adaptation in the open-vocabulary setting.
Code (0)
등록된 구현이 없습니다.
Tasks
Source-Free Domain AdaptationSemantic SegmentationResults from the Paper
| Rank | Task | Dataset | Model | Metrics |
|---|---|---|---|---|
| #17 | Semantic Segmentation | Cityscapes | VocAlign | mIoU: 6.11 |
Similar Papers 제목 키워드 기반
Non-invasive electromyographic speech neuroprosthesis: a geometric perspective
In this article, we present a high-bandwidth egocentric neuromuscular speech interface for translating silently voiced speech articulations into textand audio. Specifically, we collect electromyogram (EMG) signals from m…
Speech SynthesisThe Devil is in the Details: On the Pitfalls of Vocabulary Selection in Neural Machine Translation
Vocabulary selection, or lexical shortlisting, is a well-known technique to improve latency of Neural Machine Translation models by constraining the set of allowed output words during inference. The chosen set is typical…
Machine TranslationSentenceTranslationSpeeding Up Neural Machine Translation Decoding by Shrinking Run-time Vocabulary
We speed up Neural Machine Translation (NMT) decoding by shrinking run-time target vocabulary. We experiment with two shrinking approaches: Locality Sensitive Hashing (LSH) and word alignments. Using the latter method, w…
GPUMachine TranslationNMTTranslationEffective Cross-lingual Transfer of Neural Machine Translation Models without Shared Vocabularies
Transfer learning or multilingual model is essential for low-resource neural machine translation (NMT), but the applicability is limited to cognate languages by sharing their vocabularies. This paper shows effective tech…
Cross-Lingual TransferLow Resource Neural Machine TranslationLow-Resource Neural Machine TranslationMachine Translation+3Lost in Translation: Analysis of Information Loss During Machine Translation Between Polysynthetic and Fusional Languages
Machine translation from polysynthetic to fusional languages is a challenging task, which gets further complicated by the limited amount of parallel text available. Thus, translation performance is far from the state of …
Machine TranslationTranslation