paper-with-me

Papers

MES-RAG: Bringing Multi-modal, Entity-Storage, and Secure Enhancements to RAG

2025-03-17 · Pingyu Wu, Daiheng Gao, Jing Tang, Huimin Chen, Wenbo Zhou, Weiming Zhang, Nenghai Yu

Retrieval-Augmented Generation (RAG) improves Large Language Models (LLMs) by using external knowledge, but it struggles with precise entity information retrieval. In this paper, we proposed MES-RAG framework, which enhances entity-specific query handling and provides accurate, secure, and consistent responses. MES-RAG introduces proactive security measures that ensure system integrity by applying protections prior to data access. Additionally, the system supports real-time multi-modal outputs, including text, images, audio, and video, seamlessly integrating into existing RAG architectures. Experimental results demonstrate that MES-RAG significantly improves both accuracy and recall, highlighting its effectiveness in advancing the security and utility of question-answering, increasing accuracy to 0.83 (+0.25) on targeted task. Our code and data are available at https://github.com/wpydcr/MES-RAG.

📄 PDF Abstract BibTeX arXiv:2503.13563

Code (1)

wpydcr/mes-rag 공식 구현

Tasks

Information RetrievalQuestion AnsweringRAGRetrievalRetrieval-augmented Generation

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Attention 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Residual Connection 설명 없음
Linear Warmup With Linear Decay Linear Warmup With Linear Decay is a learning rate schedule in which we increase the learning rate linearly for $n$ updates and then linearly decay afterwards.
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Adam 설명 없음
WordPiece 설명 없음

Similar Papers 제목 키워드 기반

TAPESTRY: A Blockchain based Service for Trusted Interaction Online

2019-05-15 · Yifan Yang, Daniel Cooper, John Collomosse, Constantin C. Drăgan 외

We present a novel blockchain based service for proving the provenance of online digital identity, exposed as an assistive tool to help non-expert users make better decisions about whom to trust online. Our service harne…

Privacy Preserving

Trustera: A Live Conversation Redaction System

2023-03-16 · Evandro Gouvêa, Ali Dadgar, Shahab Jalalvand, Rathi Chengalvarayan 외

Trustera, the first functional system that redacts personally identifiable information (PII) in real-time spoken conversations to remove agents' need to hear sensitive information while preserving the naturalness of live…

Automatic Speech RecognitionNatural Language Understandingspeech-recognitionSpeech Recognition

BCGS: Blockchain-assisted privacy-preserving cross-domain authentication for VANETs

2023-03-15 · Vehicular Communications 2023 3 · Biwen Chen, Zhongming Wang, Tao Xiang, Jiyun Yang 외

Vehicular Ad-Hoc Networks (VANETs) have significantly enhanced driving safety and comfort by leveraging vehicular wireless communication technology. Secure authentication among vehicles in VANETs is an important requirem…

Privacy Preserving

Privacy-Enhanced Data Sharing Systems from Hierarchical ID-Based Puncturable Functional Encryption with Inner Product Predicates

2024-09-28 · IET Information Security 2024 9 · Cheng-Yi Lee, Zi-Yuan Liu, Masahiro Mambo, Raylin Tso

The emergence of cloud computing enables users to upload data to remote clouds and compute them. This drastically reduces computing and storage costs for users. Considering secure computing for multilevel users in enterp…

Cloud Computing

Seeing is Deceiving: Exploitation of Visual Pathways in Multi-Modal Language Models

2024-11-07 · Pete Janowczyk, Linda Laurier, Ave Giulietta, Arlo Octavia 외

Multi-Modal Language Models (MLLMs) have transformed artificial intelligence by combining visual and text data, making applications like image captioning, visual question answering, and multi-modal content creation possi…

Adversarial AttackImage CaptioningQuestion AnsweringVisual Question Answering