paper-with-me

Retrieval

11개 벤치마크 · 논문 14,297편 · 이 태스크의 논문 보기 →

Benchmarks

Natural Questions

결과 5개

HotpotQA

결과 4개

PubMedQA

결과 4개

Quora Question Pairs

결과 4개

OK-VQA

결과 2개

InfoSeek

결과 1개

MVK

결과 1개

Polyvore

결과 1개

ToolLens

결과 1개

Most implemented

Papers

From Roots to Rewards: Dynamic Tree Reasoning with RL

2025-07-17 · Ahmed Bahloul, Simon Malberg

Modern language models address complex questions through chain-of-thought (CoT) reasoning (Wei et al., 2023) and retrieval augmentation (Lewis et al., 2021), yet struggle with error propagation and knowledge integration.…

Computational EfficiencyQuestion AnsweringRetrieval

HapticCap: A Multimodal Dataset and Task for Understanding User Experience of Vibration Haptic Signals

2025-07-17 · Guimin Hu, Daniel Hershcovich, Hasti Seifi

Haptic signals, from smartphone vibrations to virtual reality touch feedback, can effectively convey information and enhance realism, but designing signals that resonate meaningfully with users is challenging. To facilit…

Contrastive LearningRetrieval

A Survey of Context Engineering for Large Language Models

2025-07-17 · Lingrui Mei, Jiayu Yao, Yuyao Ge, Yiwei Wang 외

The performance of Large Language Models (LLMs) is fundamentally determined by the contextual information provided during inference. This survey introduces Context Engineering, a formal discipline that transcends simple …

RAGRetrievalRetrieval-augmented GenerationSurvey

MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval

2025-07-17 · Jeong-Woo Park, Seong-Whan Lee

Composed Image Retrieval (CIR) is the task of retrieving a target image from a gallery using a composed query consisting of a reference image and a modification text. Among various CIR approaches, training-free zero-shot…

Image RetrievalRe-RankingRetrieval

Developing Visual Augmented Q&A System using Scalable Vision Embedding Retrieval & Late Interaction Re-ranker

2025-07-16 · Rachna Saxena, Abhijeet Kumar, Suresh Shanmugam

Traditional information extraction systems face challenges with text only language models as it does not consider infographics (visual elements of information) such as tables, charts, images etc. often used to convey com…

RAGRetrieval

Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos

2025-07-16 · Yuchi Ishikawa, Shota Nakada, Hokuto Munakata, Kazuhiro Saito 외

In this paper, we propose Language-Guided Contrastive Audio-Visual Masked Autoencoders (LG-CAV-MAE) to improve audio-visual representation learning. LG-CAV-MAE integrates a pretrained text encoder into contrastive audio-…

Image CaptioningRepresentation LearningRetrieval

전체 14,297편 보기 →