Embedding Meta-Textual Information for Improved Learning to Rank
Neural approaches to learning term embeddings have led to improved computation of similarity and ranking in information retrieval (IR). So far neural representation learning has not been extended to meta-textual information that is readily available for many IR tasks, for example, patent classes in prior-art retrieval, topical information in Wikipedia articles, or product categories in e-commerce data. We present a framework that learns embeddings for meta-textual categories, and optimizes a pairwise ranking objective for improved matching based on combined embeddings of textual and meta-textual information. We show considerable gains in an experimental evaluation on cross-lingual retrieval in the Wikipedia domain for three language pairs, and in the Patent domain for one language pair. Our results emphasize that the mode of combining different types of information is crucial for model improvement.
Code (0)
등록된 구현이 없습니다.
Tasks
ArticlesInformation RetrievalLearning-To-RankRepresentation LearningRetrievalSimilar Papers 제목 키워드 기반
Metadata-Driven Retrieval-Augmented Generation for Financial Question Answering
Retrieval-Augmented Generation (RAG) struggles on long, structured financial filings where relevant evidence is sparse and cross-referenced. This paper presents a systematic investigation of advanced metadata-driven Retr…
Question AnsweringA Feature Analysis for Multimodal News Retrieval
Content-based information retrieval is based on the information contained in documents rather than using metadata such as keywords. Most information retrieval methods are either based on text or image. In this paper, we …
Information RetrievalNews RetrievalRetrievalWord EmbeddingsMeta-Embedding Sentence Representation for Textual Similarity
Word embedding models are now widely used in most NLP applications. Despite their effectiveness, there is no clear evidence about the choice of the most appropriate model. It often depends on the nature of the task and o…
Question SimilaritySentenceSentence EmbeddingSentence-Embedding+1Exploiting Position and Contextual Word Embeddings for Keyphrase Extraction from Scientific Papers
Keyphrases associated with research papers provide an effective way to find useful information in the large and growing scholarly digital collections. In this paper, we present KPRank, an unsupervised graph-based algorit…
Keyphrase ExtractionPositionWord EmbeddingsScaling User Modeling: Large-scale Online User Representations for Ads Personalization in Meta
Effective user representations are pivotal in personalized advertising. However, stringent constraints on training throughput, serving latency, and memory, often limit the complexity and input feature set of online ads r…
Representation Learning