paper-with-me

홈 › Papers

Context-Enhanced Language Models for Generating Multi-Paper Citations

2024-04-22 · Avinash Anand, Kritarth Prasad, Ujjwal Goel, Mohit Gupta, Naman Lal, Astha Verma, Rajiv Ratn Shah

Citation text plays a pivotal role in elucidating the connection between scientific documents, demanding an in-depth comprehension of the cited paper. Constructing citations is often time-consuming, requiring researchers to delve into extensive literature and grapple with articulating relevant content. To address this challenge, the field of citation text generation (CTG) has emerged. However, while earlier methods have primarily centered on creating single-sentence citations, practical scenarios frequently necessitate citing multiple papers within a single paragraph. To bridge this gap, we propose a method that leverages Large Language Models (LLMs) to generate multi-citation sentences. Our approach involves a single source paper and a collection of target papers, culminating in a coherent paragraph containing multi-sentence citation text. Furthermore, we introduce a curated dataset named MCG-S2ORC, composed of English-language academic research papers in Computer Science, showcasing multiple citation instances. In our experiments, we evaluate three LLMs LLaMA, Alpaca, and Vicuna to ascertain the most effective model for this endeavor. Additionally, we exhibit enhanced performance by integrating knowledge graphs from target papers into the prompts for generating citation text. This research underscores the potential of harnessing LLMs for citation generation, opening a compelling avenue for exploring the intricate connections between scientific documents.

📄 PDF Abstract BibTeX arXiv:2404.13865

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge GraphsSentenceText Generation

Methods 이 논문이 사용한 방법론

LLaMA LLaMA is a collection of foundation language models ranging from 7B to 65B parameters. It is based on the transformer architecture with various improvements that were…

Similar Papers 제목 키워드 기반

Large Body Language Models

2024-10-21 · Saif Punjwani, Larry Heck

As virtual agents become increasingly prevalent in human-computer interaction, generating realistic and contextually appropriate gestures in real-time remains a significant challenge. While neural rendering techniques ha…

Gesture GenerationLanguage ModelingLanguage ModellingLarge Language Model+1

WisdoM: Improving Multimodal Sentiment Analysis by Fusing Contextual World Knowledge

2024-01-12 · Wenbin Wang, Liang Ding, Li Shen, Yong Luo 외

Sentiment analysis is rapidly advancing by utilizing various data modalities (e.g., text, image). However, most previous works relied on superficial information, neglecting the incorporation of contextual world knowledge…

Multimodal Sentiment AnalysisSentiment AnalysisWorld Knowledge

Building a Llama2-finetuned LLM for Odia Language Utilizing Domain Knowledge Instruction Set

2023-12-19 · Guneet Singh Kohli, Shantipriya Parida, Sambit Sekhar, Samirit Saha 외

Building LLMs for languages other than English is in great demand due to the unavailability and performance of multilingual LLMs, such as understanding the local context. The problem is critical for low-resource language…

MRRC: Multiple Role Representation Crossover Interpretation for Image Captioning With R-CNN Feature Distribution Composition (FDC)

2020-02-15 · Chiranjib Sur

While image captioning through machines requires structured learning and basis for interpretation, improvement requires multiple context understanding and processing in a meaningful way. This research will provide a nove…

DecoderImage CaptioningReinforcement LearningSentence

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection

2025-02-12 · Junjie Wu, Yumeng Fu, Nan Yu, Guohong Fu

Recent studies in Large Vision-Language Models (LVLMs) have demonstrated impressive advancements in multimodal Out-of-Context (OOC) misinformation detection, discerning whether an authentic image is wrongly used in a cla…

Instruction FollowingLanguage ModelingLanguage ModellingMisinformation+1