paper-with-me

홈 › Papers

Constructing a Visual Relationship Authenticity Dataset

2020-10-11 · Chenhui Chu, Yuto Takebayashi, Mishra Vipul, Yuta Nakashima

A visual relationship denotes a relationship between two objects in an image, which can be represented as a triplet of (subject; predicate; object). Visual relationship detection is crucial for scene understanding in images. Existing visual relationship detection datasets only contain true relationships that correctly describe the content in an image. However, distinguishing false visual relationships from true ones is also crucial for image understanding and grounded natural language processing. In this paper, we construct a visual relationship authenticity dataset, where both true and false relationships among all objects appeared in the captions in the Flickr30k entities image caption dataset are annotated. The dataset is available at https://github.com/codecreator2053/VR_ClassifiedDataset. We hope that this dataset can promote the study on both vision and language understanding.

📄 PDF Abstract BibTeX arXiv:2010.05185

Code (0)

등록된 구현이 없습니다.

Tasks

Relationship DetectionScene UnderstandingTripletVisual Relationship Detection

Similar Papers 제목 키워드 기반

Topology Imbalance and Relation Inauthenticity Aware Hierarchical Graph Attention Networks for Fake News Detection

2022-10-01 · COLING 2022 10 · Li Gao, Lingyun Song, Jie Liu, Bolin Chen 외

Fake news detection is a challenging problem due to its tremendous real-world political and social impacts. Recent fake news detection works focus on learning news features from News Propagation Graph (NPG). However, lit…

Fake News DetectionGraph AttentionRelation

Grounding and Explaining Visual Evidence for AI-Generated Image Detection in Human-Centric Scenes

2026-08-03 · Kun Guo, Yuzhou Yang, Haoyue Wang, Qichao Ying 외 arxiv

Rapid advances in image generation models call for interpretable AI-generated image detection methods that not only determine authenticity but also provide supporting visual evidence. Existing approaches may produce inco…

From Talking to Singing: A New Challenge for Audio-Visual Deepfake Detection

2026-05-27 · Ke Liu, Jiwei Wei, Wenyu Zhang, Shuchang Zhou 외 arxiv

With rapid advances in audio-visual generative models, reliable forgery detection becomes increasingly critical. Existing methods for audio-visual deepfake detection typically rely on cross-modal inconsistencies. In sing…

DeepFake Detection

LogicLens: Visual-Logical Co-Reasoning for Text-Centric Forgery Analysis

2025-12-25 · Fanwei Zeng, Changtao Miao, Jing Huang, Zhiya Tan 외 arxiv

Sophisticated text-centric forgeries, fueled by rapid AIGC advancements, pose a significant threat to societal security and information authenticity. Current methods for text-centric forgery analysis are often limited to…

GazeVaLM: A Multi-Observer Eye-Tracking Benchmark for Evaluating Clinical Realism in AI-Generated X-Rays

2026-04-13 · David Wong, Zeynep Isik, Bin Wang, Marouane Tliba 외 arxiv

We introduce GazeVaLM, a public eye-tracking dataset for studying clinical perception during chest radiograph authenticity assessment. The dataset comprises 960 gaze recordings from 16 expert radiologists interpreting 30…