paper-with-me

Papers Retrieval

“Retrieval” 태그가 달린 논문 14,297편 · 필터 해제

From Roots to Rewards: Dynamic Tree Reasoning with RL

2025-07-17 · Ahmed Bahloul, Simon Malberg

Modern language models address complex questions through chain-of-thought (CoT) reasoning (Wei et al., 2023) and retrieval augmentation (Lewis et al., 2021), yet struggle with error propagation and knowledge integration.…

Computational EfficiencyQuestion AnsweringRetrieval

HapticCap: A Multimodal Dataset and Task for Understanding User Experience of Vibration Haptic Signals

2025-07-17 · Guimin Hu, Daniel Hershcovich, Hasti Seifi

Haptic signals, from smartphone vibrations to virtual reality touch feedback, can effectively convey information and enhance realism, but designing signals that resonate meaningfully with users is challenging. To facilit…

Contrastive LearningRetrieval

A Survey of Context Engineering for Large Language Models

2025-07-17 · Lingrui Mei, Jiayu Yao, Yuyao Ge, Yiwei Wang 외

The performance of Large Language Models (LLMs) is fundamentally determined by the contextual information provided during inference. This survey introduces Context Engineering, a formal discipline that transcends simple …

RAGRetrievalRetrieval-augmented GenerationSurvey

MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval

2025-07-17 · Jeong-Woo Park, Seong-Whan Lee

Composed Image Retrieval (CIR) is the task of retrieving a target image from a gallery using a composed query consisting of a reference image and a modification text. Among various CIR approaches, training-free zero-shot…

Image RetrievalRe-RankingRetrieval

Developing Visual Augmented Q&A System using Scalable Vision Embedding Retrieval & Late Interaction Re-ranker

2025-07-16 · Rachna Saxena, Abhijeet Kumar, Suresh Shanmugam

Traditional information extraction systems face challenges with text only language models as it does not consider infographics (visual elements of information) such as tables, charts, images etc. often used to convey com…

RAGRetrieval

Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos

2025-07-16 · Yuchi Ishikawa, Shota Nakada, Hokuto Munakata, Kazuhiro Saito 외

In this paper, we propose Language-Guided Contrastive Audio-Visual Masked Autoencoders (LG-CAV-MAE) to improve audio-visual representation learning. LG-CAV-MAE integrates a pretrained text encoder into contrastive audio-…

Image CaptioningRepresentation LearningRetrieval

Context-Aware Search and Retrieval Over Erasure Channels

2025-07-16 · Sara Ghasvarianjahromi, Yauhen Yakimenka, Jörg Kliewer

This paper introduces and analyzes a search and retrieval model that adopts key semantic communication principles from retrieval-augmented generation. We specifically present an information-theoretic analysis of a remote…

DecoderRetrievalRetrieval-augmented GenerationSemantic Communication

Seq vs Seq: An Open Suite of Paired Encoders and Decoders

2025-07-15 · Orion Weller, Kathryn Ricci, Marc Marone, Antoine Chaffin 외

The large language model (LLM) community focuses almost exclusively on decoder-only language models, since they are easier to use for text generation. However, a large subset of the community still uses encoder-only mode…

DecoderLarge Language ModelRetrievalText Generation

From Chaos to Automation: Enabling the Use of Unstructured Data for Robotic Process Automation

2025-07-15 · Kelly Kurowski, Xixi Lu, Hajo A. Reijers

The growing volume of unstructured data within organizations poses significant challenges for data analysis and process automation. Unstructured data, which lacks a predefined format, encompasses various forms such as em…

Information RetrievalRetrieval

Kodezi Chronos: A Debugging-First Language Model for Repository-Scale, Memory-Driven Code Understanding

2025-07-14 · Ishraq Khan, Assad Chowdary, Sharoz Haseeb, Urvish Patel

Large Language Models (LLMs) have advanced code generation and software automation, but are fundamentally constrained by limited inference-time context and lack of explicit code structure reasoning. We introduce Kodezi C…

Code GenerationLanguage ModelingLanguage ModellingRetrieval

RadiomicsRetrieval: A Customizable Framework for Medical Image Retrieval Using Radiomics Features

2025-07-11 · Inye Na, Nejung Rue, Jiwon Chung, HyunJin Park

Medical image retrieval is a valuable field for supporting clinical decision-making, yet current methods primarily support 2D images and require fully annotated queries, limiting clinical flexibility. To address this, we…

Contrastive LearningImage RetrievalMedical Image RetrievalRetrieval+1

CLI-RAG: A Retrieval-Augmented Framework for Clinically Structured and Context Aware Text Generation with LLMs

2025-07-09 · Garapati Keerthana, Manik Gupta

Large language models (LLMs), including zero-shot and few-shot paradigms, have shown promising capabilities in clinical text generation. However, real-world applications face two key challenges: (1) patient data is highl…

ChunkingRAGRetrievalRetrieval-augmented Generation+1

Multi-Agent Retrieval-Augmented Framework for Evidence-Based Counterspeech Against Health Misinformation

2025-07-09 · Anirban Saha Anik, Xiaoying Song, Elliott Wang, Bryan Wang 외

Large language models (LLMs) incorporated with Retrieval-Augmented Generation (RAG) have demonstrated powerful capabilities in generating counterspeech against misinformation. However, current studies rely on limited evi…

InformativenessMisinformationRAGRetrieval+1

Temporal Information Retrieval via Time-Specifier Model Merging

2025-07-09 · SeungYoon Han, Taeho Hwang, Sukmin Cho, Soyeong Jeong 외

The rapid expansion of digital information and knowledge across structured and unstructured sources has heightened the importance of Information Retrieval (IR). While dense retrieval methods have substantially improved s…

Information RetrievalmodelRetrieval

Orchestrator-Agent Trust: A Modular Agentic AI Visual Classification System with Trust-Aware Orchestration and RAG-Based Reasoning

2025-07-09 · Konstantinos I. Roumeliotis, Ranjan Sapkota, Manoj Karkee, Nikolaos D. Tselikas

Modern Artificial Intelligence (AI) increasingly relies on multi-agent architectures that blend visual and language understanding. Yet, a pressing challenge remains: How can we trust these agents especially in zero-shot …

BenchmarkingImage RetrievalOptical Character Recognition (OCR)RAG+3

SARA: Selective and Adaptive Retrieval-augmented Generation with Context Compression

2025-07-08 · Yiqiao Jin, Kartik Sharma, Vineeth Rakesh, Yingtong Dou 외

Retrieval-augmented Generation (RAG) extends large language models (LLMs) with external knowledge but faces key challenges: restricted effective context length and redundancy in retrieved documents. Pure compression-base…

Evidence SelectionRAGRerankingRetrieval+4

Differential Mamba

2025-07-08 · Nadav Schneider, Itamar Zimerman, Eliya Nachmani

Sequence models like Transformers and RNNs often overallocate attention to irrelevant context, leading to noisy intermediate representations. This degrades LLM capabilities by promoting hallucinations, weakening long-ran…

Language ModelingLanguage ModellingMambaRetrieval

Semantic Certainty Assessment in Vector Retrieval Systems: A Novel Framework for Embedding Quality Evaluation

2025-07-08 · Y. Du

Vector retrieval systems exhibit significant performance variance across queries due to heterogeneous embedding quality. We propose a lightweight framework for predicting retrieval performance at the query level by combi…

Data AugmentationQuantizationRetrieval

Automatic Synthesis of High-Quality Triplet Data for Composed Image Retrieval

2025-07-08 · Haiwen Li, Delong Liu, Zhaohui Hou, Zhicheng Zhao 외

As a challenging vision-language (VL) task, Composed Image Retrieval (CIR) aims to retrieve target images using multimodal (image+text) queries. Although many existing CIR methods have attained promising performance, the…

Image RetrievalLarge Language ModelRetrievalTriplet

An analysis of vision-language models for fabric retrieval

2025-07-07 · Francesco Giuliari, Asif Khan Pattan, Mohamed Lamine Mekhalfi, Fabio Poiesi

Effective cross-modal retrieval is essential for applications like information retrieval and recommendation systems, particularly in specialized domains such as manufacturing, where product information often consists of …

AttributeCross-Modal RetrievalImage RetrievalInformation Retrieval+3
1–20 / 14,297 다음 →