paper-with-me

Papers

EXIF as Language: Learning Cross-Modal Associations Between Images and Camera Metadata

2023-01-11 · CVPR 2023 1 · Chenhao Zheng, Ayush Shrivastava, Andrew Owens

We learn a visual representation that captures information about the camera that recorded a given photo. To do this, we train a multimodal embedding between image patches and the EXIF metadata that cameras automatically insert into image files. Our model represents this metadata by simply converting it to text and then processing it with a transformer. The features that we learn significantly outperform other self-supervised and supervised features on downstream image forensics and calibration tasks. In particular, we successfully localize spliced image regions "zero shot" by clustering the visual embeddings for all of the patches within an image.

📄 PDF Abstract BibTeX arXiv:2301.04647

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringImage Forensics

Similar Papers 제목 키워드 기반

Mapping 'when'-clauses in Latin American and Caribbean languages: an experiment in subtoken-based typology

2024-04-28 · Nilo Pedrazzini

Languages can encode temporal subordination lexically, via subordinating conjunctions, and morphologically, by marking the relation on the predicate. Systematic cross-linguistic variation among the former can be studied …

Universal Conceptual Structure in Neural Translation: Probing NLLB-200's Multilingual Geometry

2026-02-27 · Kyle Elliott Mathewson arxiv

Do neural machine translation models learn language-universal conceptual representations, or do they merely cluster languages by surface similarity? We investigate this question by probing the representation geometry of …

Machine Translation

Investigating Lexical Change through Cross-Linguistic Colexification Patterns

2025-10-15 · Kim Gfeller, Sabine Stoll, Chundra Cathcart, Paul Widmer arxiv

One of the most intriguing features of language is its constant change, with ongoing shifts in how meaning is expressed. Despite decades of research, the factors that determine how and why meanings evolve remain only par…

Conceptual similarity and communicative need shape colexification: an experimental study

2021-03-19 · Andres Karjus, Richard A. Blythe, Simon Kirby, Tianyu Wang 외

Colexification refers to the phenomenon of multiple meanings sharing one word in a language. Cross-linguistic lexification patterns have been shown to be largely predictable, as similar concepts are often colexified. We …

What does Kiki look like? Cross-modal associations between speech sounds and visual shapes in vision-and-language models

2024-07-25 · Tessa Verhoef, Kiana Shahrasbi, Tom Kouwenhoven

Humans have clear cross-modal preferences when matching certain novel words to visual shapes. Evidence suggests that these preferences play a prominent role in our linguistic processing, language learning, and the origin…