Retrieval
11개 벤치마크 · 논문 14,297편 · 이 태스크의 논문 보기 →
Benchmarks
Natural Questions
HotpotQA
PubMedQA
Quora Question Pairs
OK-VQA
InfoSeek
MVK
Polyvore
ToolLens
Most implemented
Beyond Part Models: Person Retrieval with Refined Part Pooling (and a Strong Convolutional Baseline)
Modeling Relational Data with Graph Convolutional Networks
TransferTransfo: A Transfer Learning Approach for Neural Network Based Conversational Agents
Dense Passage Retrieval for Open-Domain Question Answering
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
Circle Loss: A Unified Perspective of Pair Similarity Optimization
Papers
From Roots to Rewards: Dynamic Tree Reasoning with RL
Modern language models address complex questions through chain-of-thought (CoT) reasoning (Wei et al., 2023) and retrieval augmentation (Lewis et al., 2021), yet struggle with error propagation and knowledge integration.…
Computational EfficiencyQuestion AnsweringRetrievalHapticCap: A Multimodal Dataset and Task for Understanding User Experience of Vibration Haptic Signals
Haptic signals, from smartphone vibrations to virtual reality touch feedback, can effectively convey information and enhance realism, but designing signals that resonate meaningfully with users is challenging. To facilit…
Contrastive LearningRetrievalA Survey of Context Engineering for Large Language Models
The performance of Large Language Models (LLMs) is fundamentally determined by the contextual information provided during inference. This survey introduces Context Engineering, a formal discipline that transcends simple …
RAGRetrievalRetrieval-augmented GenerationSurveyMCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval
Composed Image Retrieval (CIR) is the task of retrieving a target image from a gallery using a composed query consisting of a reference image and a modification text. Among various CIR approaches, training-free zero-shot…
Image RetrievalRe-RankingRetrievalDeveloping Visual Augmented Q&A System using Scalable Vision Embedding Retrieval & Late Interaction Re-ranker
Traditional information extraction systems face challenges with text only language models as it does not consider infographics (visual elements of information) such as tables, charts, images etc. often used to convey com…
RAGRetrievalLanguage-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos
In this paper, we propose Language-Guided Contrastive Audio-Visual Masked Autoencoders (LG-CAV-MAE) to improve audio-visual representation learning. LG-CAV-MAE integrates a pretrained text encoder into contrastive audio-…
Image CaptioningRepresentation LearningRetrieval