paper-with-me

홈 › Papers

Prompt me a Dataset: An investigation of text-image prompting for historical image dataset creation using foundation models

2023-09-04 · Hassan El-Hajj, Matteo Valleriani

In this paper, we present a pipeline for image extraction from historical documents using foundation models, and evaluate text-image prompts and their effectiveness on humanities datasets of varying levels of complexity. The motivation for this approach stems from the high interest of historians in visual elements printed alongside historical texts on the one hand, and from the relative lack of well-annotated datasets within the humanities when compared to other domains. We propose a sequential approach that relies on GroundDINO and Meta's Segment-Anything-Model (SAM) to retrieve a significant portion of visual data from historical documents that can then be used for downstream development tasks and dataset creation, as well as evaluate the effect of different linguistic prompts on the resulting detections.

📄 PDF Abstract BibTeX arXiv:2309.01674

Code (1)

hassanhajj910/prompt-me-a-dataset 공식 구현 pytorch

Similar Papers 제목 키워드 기반

LoR-VP: Low-Rank Visual Prompting for Efficient Vision Model Adaptation

2025-02-02 · Can Jin, Ying Li, Mingyu Zhao, Shiyu Zhao 외

Visual prompting has gained popularity as a method for adapting pre-trained models to specific tasks, particularly in the realm of parameter-efficient tuning. However, existing visual prompting techniques often pad the p…

Inductive BiasVisual Prompting

Chain-of-Thought Augmentation with Logit Contrast for Enhanced Reasoning in Language Models

2024-07-04 · Jay Shim, Grant Kruttschnitt, Alyssa Ma, Daniel Kim 외

Rapidly increasing model scales coupled with steering methods such as chain-of-thought prompting have led to drastic improvements in language model reasoning. At the same time, models struggle with compositional generali…

Language ModelingLanguage Modelling

LLMs for Low-Resource Dialect Translation Using Context-Aware Prompting: A Case Study on Sylheti

2025-11-24 · Tabia Tanzin Prama, Christopher M. Danforth, Peter Sheridan Dodds arxiv

Large Language Models (LLMs) have demonstrated strong translation abilities through prompting, even without task-specific training. However, their effectiveness in dialectal and low-resource contexts remains underexplore…

Machine Translation

Prompting Techniques for Secure Code Generation: A Systematic Investigation

2024-07-09 · Catherine Tony, Nicolás E. Díaz Ferreyra, Markus Mutas, Salem Dhiff 외

Large Language Models (LLMs) are gaining momentum in software development with prompt-driven programming enabling developers to create code from natural language (NL) instructions. However, studies have questioned their …

Code GenerationSystematic Literature Review

Decomposed Prompting: Unveiling Multilingual Linguistic Structure Knowledge in English-Centric Large Language Models

2024-02-28 · Ercong Nie, Shuzhou Yuan, Bolei Ma, Helmut Schmid 외

Despite the predominance of English in their training data, English-centric Large Language Models (LLMs) like GPT-3 and LLaMA display a remarkable ability to perform multilingual tasks, raising questions about the depth …

Part-Of-Speech TaggingSentence