paper-with-me

홈 › Papers

Visual Objectification in Films: Towards a New AI Task for Video Interpretation

2024-01-24 · CVPR 2024 1 · Julie Tores, Lucile Sassatelli, Hui-Yin Wu, Clement Bergman, Lea Andolfi, Victor Ecrement, Frederic Precioso, Thierry Devars, Magali Guaresi, Virginie Julliard, Sarah Lecossais

In film gender studies, the concept of 'male gaze' refers to the way the characters are portrayed on-screen as objects of desire rather than subjects. In this article, we introduce a novel video-interpretation task, to detect character objectification in films. The purpose is to reveal and quantify the usage of complex temporal patterns operated in cinema to produce the cognitive perception of objectification. We introduce the ObyGaze12 dataset, made of 1914 movie clips densely annotated by experts for objectification concepts identified in film studies and psychology. We evaluate recent vision models, show the feasibility of the task and where the challenges remain with concept bottleneck models. Our new dataset and code are made available to the community.

📄 PDF Abstract BibTeX arXiv:2401.13296

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MObyGaze: a film dataset of multimodal objectification densely annotated by experts

2025-05-28 · Julie Tores, Elisa Ancarani, Lucile Sassatelli, Hui-Yin Wu 외 arxiv

Characterizing and quantifying gender representation disparities in audiovisual storytelling contents is necessary to grasp how stereotypes may perpetuate on screen. In this article, we consider the high-level construct …

Reflecting the Male Gaze: Quantifying Female Objectification in 19th and 20th Century Novels

2024-03-25 · Kexin Luo, Yue Mao, Bei Zhang, Sophie Hao

Inspired by the concept of the male gaze (Mulvey, 1975) in literature and media studies, this paper proposes a framework for analyzing gender bias in terms of female objectification: the extent to which a text portrays f…

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?

2025-05-27 · Junhao Cheng, Yuying Ge, Teng Wang, Yixiao Ge 외

Recent advances in CoT reasoning and RL post-training have been reported to enhance video reasoning capabilities of MLLMs. This progress naturally raises a question: can these models perform complex video reasoning in a …

Multimodal Reasoning

Multimodal Sexism Identification and Characterization using Large Language Models and Gradient Boosting

2026-06-04 · Kyriakos Chaviaras, Maria Lymperaiou, Athanasios Voulodimos arxiv

We present the AILS-NTUA submission to the EXIST 2026 Lab at CLEF, addressing multimodal sexism identification and characterization in memes (Task 2) and short-form videos (Task 3). Our system follows a feature-engineere…

Feature Engineering

Bridging Your Imagination with Audio-Video Generation via a Unified Director

2025-12-29 · Jiaxu Zhang, Tianshu Hu, Yuan Zhang, Zenan Li 외 arxiv

Existing AI-driven video creation systems typically treat script drafting and key-shot design as two disjoint tasks: the former relies on large language models, while the latter depends on image generation models. We arg…

Logical ReasoningVideo GenerationImage Generation