paper-with-me

Papers

Augmenting a Large Language Model with a Combination of Text and Visual Data for Conversational Visualization of Global Geospatial Data

2025-01-16 · Omar Mena, Alexandre Kouyoumdjian, Lonni Besançon, Michael Gleicher, Ivan Viola, Anders Ynnerman

We present a method for augmenting a Large Language Model (LLM) with a combination of text and visual data to enable accurate question answering in visualization of scientific data, making conversational visualization possible. LLMs struggle with tasks like visual data interaction, as they lack contextual visual information. We address this problem by merging a text description of a visualization and dataset with snapshots of the visualization. We extract their essential features into a structured text file, highly compact, yet descriptive enough to appropriately augment the LLM with contextual information, without any fine-tuning. This approach can be applied to any visualization that is already finally rendered, as long as it is associated with some textual description.

📄 PDF Abstract BibTeX arXiv:2501.09521

Code (0)

등록된 구현이 없습니다.

Tasks

Data InteractionDescriptiveLanguage ModelingLanguage ModellingLarge Language ModelQuestion Answering

Similar Papers 제목 키워드 기반

UAlberta at SemEval-2023 Task 1: Context Augmentation and Translation for Multilingual Visual Word Sense Disambiguation

2023-06-24 · Michael Ogezi, Bradley Hauer, Talgat Omarov, Ning Shi 외

We describe the systems of the University of Alberta team for the SemEval-2023 Visual Word Sense Disambiguation (V-WSD) Task. We present a novel algorithm that leverages glosses retrieved from BabelNet, in combination wi…

Image GenerationImage SegmentationLanguage ModelingLanguage Modelling+2

MAGMA -- Multimodal Augmentation of Generative Models through Adapter-based Finetuning

2021-12-09 · Constantin Eichenberg, Sidney Black, Samuel Weinbach, Letitia Parcalabescu 외

Large-scale pretraining is fast becoming the norm in Vision-Language (VL) modeling. However, prevailing VL approaches are limited by the requirement for labeled data and the use of complex multi-step pretraining objectiv…

In-Context LearningLanguage ModelingLanguage Modelling

Examining the Effects of Language-and-Vision Data Augmentation for Generation of Descriptions of Human Faces

2022-06-01 · PVLAM (LREC) 2022 6 · Nikolai Ilinykh, Rafal Černiavski, Eva Elžbieta Sventickaitė, Viktorija Buzaitė 외

We investigate how different augmentation techniques on both textual and visual representations affect the performance of the face description generation model. Specifically, we provide the model with either original ima…

Caption GenerationData AugmentationImage Captioning

Make-it-Real: Unleashing Large Multimodal Model for Painting 3D Objects with Realistic Materials

2024-04-25 · Ye Fang, Zeyi Sun, Tong Wu, Jiaqi Wang 외

Physically realistic materials are pivotal in augmenting the realism of 3D assets across various applications and lighting conditions. However, existing 3D assets and generative models often lack authentic material prope…

Augmenting Knowledge Graph Hierarchies Using Neural Transformers

2024-04-11 · Sanat Sharma, Mayank Poddar, Jayant Kumar, Kosta Blank 외

Knowledge graphs are useful tools to organize, recommend and sort data. Hierarchies in knowledge graphs provide significant benefit in improving understanding and compartmentalization of the data within a knowledge graph…

Knowledge Graphs