paper-with-me

홈 › Papers

Investigating the Effects of Sparse Attention on Cross-Encoders

2023-12-29 · Ferdinand Schlatt, Maik Fröbe, Matthias Hagen

Cross-encoders are effective passage and document re-rankers but less efficient than other neural or classic retrieval models. A few previous studies have applied windowed self-attention to make cross-encoders more efficient. However, these studies did not investigate the potential and limits of different attention patterns or window sizes. We close this gap and systematically analyze how token interactions can be reduced without harming the re-ranking effectiveness. Experimenting with asymmetric attention and different window sizes, we find that the query tokens do not need to attend to the passage or document tokens for effective re-ranking and that very small window sizes suffice. In our experiments, even windows of 4 tokens still yield effectiveness on par with previous cross-encoders while reducing the memory requirements by at least 22% / 59% and being 1% / 43% faster at inference time for passages / documents.

📄 PDF Abstract BibTeX arXiv:2312.17649

Code (1)

webis-de/ecir-24 공식 구현

Tasks

Re-RankingRetrieval

Similar Papers 제목 키워드 기반

Hybrid Autoregressive Inference for Scalable Multi-hop Explanation Regeneration

2021-07-25 · Marco Valentino, Mokanarangan Thayaparan, Deborah Ferreira, André Freitas

Regenerating natural language explanations in the scientific domain has been proposed as a benchmark to evaluate complex multi-hop and explainable inference. In this context, large language models can achieve state-of-th…

Multi-hop Question AnsweringNatural Language InferenceQuestion Answering

Understanding Refusal in Language Models with Sparse Autoencoders

2025-05-29 · Wei Jie Yeo, Nirmalendu Prakash, Clement Neo, Roy Ka-Wei Lee 외

Refusal is a key safety behavior in aligned language models, yet the internal mechanisms driving refusals remain opaque. In this work, we conduct a mechanistic study of refusal in instruction-tuned LLMs using sparse auto…

Causal Concept Graphs in LLM Latent Space for Stepwise Reasoning

2026-03-11 · Md Muntaqim Meherab, Noor Islam S. Mohammad, Faiza Feroz arxiv

Sparse autoencoders can localize where concepts live in language models, but not how they interact during multi-step reasoning. We propose Causal Concept Graphs (CCG): a directed acyclic graph over sparse, interpretable …

Steering Language Model Refusal with Sparse Autoencoders

2024-11-18 · Kyle O'Brien, David Majercak, Xavier Fernandes, Richard Edgar 외

Responsible practices for deploying language models include guiding models to recognize and refuse answering prompts that are considered unsafe, while complying with safe prompts. Achieving such behavior typically requir…

Language ModelingLanguage Modellingmodel

Encoders Help You Disambiguate Word Senses in Neural Machine Translation

2019-08-30 · IJCNLP 2019 11 · Gongbo Tang, Rico Sennrich, Joakim Nivre

Neural machine translation (NMT) has achieved new state-of-the-art performance in translating ambiguous words. However, it is still unclear which component dominates the process of disambiguation. In this paper, we explo…

DecoderMachine TranslationNMTTranslation+1