paper-with-me

홈 › Papers

Leveraging Estimated Transferability Over Human Intuition for Model Selection in Text Ranking

2024-09-24 · Jun Bai, Zhuofan Chen, Zhenzi Li, Hanhua Hong, Jianfei Zhang, Chen Li, Chenghua Lin, Wenge Rong

Text ranking has witnessed significant advancements, attributed to the utilization of dual-encoder enhanced by Pre-trained Language Models (PLMs). Given the proliferation of available PLMs, selecting the most effective one for a given dataset has become a non-trivial challenge. As a promising alternative to human intuition and brute-force fine-tuning, Transferability Estimation (TE) has emerged as an effective approach to model selection. However, current TE methods are primarily designed for classification tasks, and their estimated transferability may not align well with the objectives of text ranking. To address this challenge, we propose to compute the expected rank as transferability, explicitly reflecting the model's ranking capability. Furthermore, to mitigate anisotropy and incorporate training dynamics, we adaptively scale isotropic sentence embeddings to yield an accurate expected rank score. Our resulting method, Adaptive Ranking Transferability (AiRTran), can effectively capture subtle differences between models. On challenging model selection scenarios across various text ranking datasets, it demonstrates significant improvements over previous classification-oriented TE methods, human intuition, and ChatGPT with minor time consumption.

📄 PDF Abstract BibTeX arXiv:2409.16198

Code (1)

ba1jun/model-selection-airtran 공식 구현 pytorch

Tasks

Model SelectionSentenceSentence Embeddings

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Evidence > Intuition: Transferability Estimation for Encoder Selection

2022-10-20 · Elisa Bassignana, Max Müller-Eberstein, Mike Zhang, Barbara Plank

With the increase in availability of large pre-trained language models (LMs) in Natural Language Processing (NLP), it becomes critical to assess their fit for a specific target task a priori - as fine-tuning the entire s…

Structured Prediction

Mimicking Human Intuition: Cognitive Belief-Driven Q-Learning

2024-10-02 · Xingrui Gu, Guanren Qiao, Chuyi Jiang, Tianqing Xia 외

Reinforcement learning encounters challenges in various environments related to robustness and explainability. Traditional Q-learning algorithms cannot effectively make decisions and utilize the historical learning exper…

Decision MakingQ-Learning

Reducing Adversarial Example Transferability Using Gradient Regularization

2019-04-16 · George Adam, Petr Smirnov, Benjamin Haibe-Kains, Anna Goldenberg

Deep learning algorithms have increasingly been shown to lack robustness to simple adversarial examples (AdvX). An equally troubling observation is that these adversarial examples transfer between different architectures…

MvP: Multi-view Prompting Improves Aspect Sentiment Tuple Prediction

2023-05-22 · Zhibin Gou, Qingyan Guo, Yujiu Yang

Generative methods greatly promote aspect-based sentiment analysis via generating a sequence of sentiment elements in a specified format. However, existing studies usually predict sentiment elements in a fixed order, whi…

Aspect-Based Sentiment AnalysisAspect-Based Sentiment Analysis (ABSA)Aspect Category DetectionAspect-Category-Opinion-Sentiment Quadruple Extraction+12

Crafting Adversarial Examples for Neural Machine Translation

2021-08-01 · ACL 2021 5 · Xinze Zhang, Junzhe Zhang, Zhenhua Chen, Kun He

Effective adversary generation for neural machine translation (NMT) is a crucial prerequisite for building robust machine translation systems. In this work, we investigate veritable evaluations of NMT adversarial attacks…

Machine TranslationNMTTranslationvalid