paper-with-me

Papers

Unveiling Key Aspects of Fine-Tuning in Sentence Embeddings: A Representation Rank Analysis

2024-05-18 · Euna Jung, Jaeill Kim, Jungmin Ko, Jinwoo Park, Wonjong Rhee

The latest advancements in unsupervised learning of sentence embeddings predominantly involve employing contrastive learning-based (CL-based) fine-tuning over pre-trained language models. In this study, we analyze the latest sentence embedding methods by adopting representation rank as the primary tool of analysis. We first define Phase 1 and Phase 2 of fine-tuning based on when representation rank peaks. Utilizing these phases, we conduct a thorough analysis and obtain essential findings across key aspects, including alignment and uniformity, linguistic abilities, and correlation between performance and rank. For instance, we find that the dynamics of the key aspects can undergo significant changes as fine-tuning transitions from Phase 1 to Phase 2. Based on these findings, we experiment with a rank reduction (RR) strategy that facilitates rapid and stable fine-tuning of the latest CL-based methods. Through empirical investigations, we showcase the efficacy of RR in enhancing the performance and stability of five state-of-the-art sentence embedding methods.

📄 PDF Abstract BibTeX arXiv:2405.11297

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningSentenceSentence EmbeddingSentence-EmbeddingSentence Embeddings

Similar Papers 제목 키워드 기반

Classification and Clustering of Sentence-Level Embeddings of Scientific Articles Generated by Contrastive Learning

2024-03-30 · Gustavo Bartz Guedes, Ana Estela Antunes da Silva

Scientific articles are long text documents organized into sections, each describing aspects of the research. Analyzing scientific production has become progressively challenging due to the increase in the number of avai…

ArticlesClusteringContrastive LearningSentence+2

AspectCSE: Sentence Embeddings for Aspect-based Semantic Textual Similarity Using Contrastive Learning and Structured Knowledge

2023-07-15 · Tim Schopf, Emanuel Gerber, Malte Ostendorff, Florian Matthes

Generic sentence embeddings provide a coarse-grained approximation of semantic textual similarity but ignore specific aspects that make texts similar. Conversely, aspect-based sentence embeddings provide similarities bet…

Contrastive LearningInformation RetrievalRetrievalSemantic Textual Similarity+4

Intrinsic vs. Extrinsic Evaluation of Czech Sentence Embeddings: Semantic Relevance Doesn't Help with MT Evaluation

2025-06-25 · Petra Barančíková, Ondřej Bojar

In this paper, we compare Czech-specific and multilingual sentence embedding models through intrinsic and extrinsic evaluation paradigms. For intrinsic evaluation, we employ Costra, a complex sentence transformation data…

Machine TranslationSemantic SimilaritySemantic Textual SimilaritySentence+5

Efficient Domain Adaptation of Sentence Embeddings Using Adapters

2023-07-06 · Tim Schopf, Dennis N. Schneider, Florian Matthes

Sentence embeddings enable us to capture the semantic similarity of short texts. Most sentence embedding models are trained for general semantic textual similarity tasks. Therefore, to use sentence embeddings in a partic…

Domain AdaptationSemantic SimilaritySemantic Textual SimilaritySentence+4

Meta-Task Prompting Elicits Embeddings from Large Language Models

2024-02-28 · Yibin Lei, Di wu, Tianyi Zhou, Tao Shen 외

We introduce a new unsupervised text embedding method, Meta-Task Prompting with Explicit One-Word Limitation (MetaEOL), for generating high-quality sentence embeddings from Large Language Models (LLMs) without the need f…

Semantic Textual SimilaritySentenceSentence EmbeddingsSTS