paper-with-me

홈 › Papers

What Can We Learn from Harry Potter? An Exploratory Study of Visual Representation Learning from Atypical Videos

2025-08-29 · Qiyue Sun, Qiming Huang, Yang Yang, Hongjun Wang, Jianbo Jiao arxiv

Humans usually show exceptional generalisation and discovery ability in the open world, when being shown uncommon new concepts. Whereas most existing studies in the literature focus on common typical data from closed sets, open-world novel discovery is under-explored in videos. In this paper, we are interested in asking: What if atypical unusual videos are exposed in the learning process? To this end, we collect a new video dataset consisting of various types of unusual atypical data (e.g., sci-fi, animation, etc.). To study how such atypical data may benefit open-world learning, we feed them into the model training process for representation learning. Focusing on three key tasks in open-world learning: out-of-distribution (OOD) detection, novel category discovery (NCD), and zero-shot action recognition (ZSAR), we found that even straightforward learning approaches with atypical data consistently improve performance across various settings. Furthermore, we found that increasing the categorical diversity of the atypical samples further boosts OOD detection performance. Additionally, in the NCD task, using a smaller yet more semantically diverse set of atypical samples leads to better performance compared to using a larger but more typical dataset. In the ZSAR setting, the semantic diversity of atypical videos helps the model generalise better to unseen action classes. These observations in our extensive experimental evaluations reveal the benefits of atypical videos for visual representation learning in the open world, together with the newly proposed dataset, encouraging further studies in this direction. The project page is at: https://julysun98.github.io/atypical_dataset.

📄 PDF Abstract BibTeX arXiv:2508.21770

Code (0)

등록된 구현이 없습니다.

Tasks

Zero-Shot Action RecognitionRepresentation Learning

Similar Papers 제목 키워드 기반

The Boy Who Survived: Removing Harry Potter from an LLM is harder than reported

2024-03-06 · Adam Shostack

Recent work arXiv.2310.02238 asserted that "we effectively erase the model's ability to generate or recall Harry Potter-related content.'' This claim is shown to be overbroad. A small experiment of less than a dozen tria…

Potterian Economics

2022-08-06 · Daniel Levy, Avichai Snir

Recent studies in psychology and neuroscience offer systematic evidence that fictional works exert a surprisingly strong influence on readers and have the power to shape their opinions and worldviews. Building on these f…

Harry Potter and the Action Prediction Challenge from Natural Language

2019-05-27 · NAACL 2019 6 · David Vilares, Carlos Gómez-Rodríguez

We explore the challenge of action prediction from textual descriptions of scenes, a testbed to approximate whether text inference can be used to predict upcoming actions. As a case of study, we consider the world of the…

Large Language Models Meet Harry Potter: A Bilingual Dataset for Aligning Dialogue Agents with Characters

2022-11-13 · Nuo Chen, Yan Wang, Haiyun Jiang, Deng Cai 외

In recent years, Dialogue-style Large Language Models (LLMs) such as ChatGPT and GPT4 have demonstrated immense potential in constructing open-domain dialogue agents. However, aligning these agents with specific characte…

Dialogue GenerationIn-Context LearningPersona Dialogue in StoryRetrieval

Harry Potter is Still Here! Probing Knowledge Leakage in Targeted Unlearned Large Language Models via Automated Adversarial Prompting

2025-05-22 · Bang Trinh Tran To, Thai Le

This work presents LURK (Latent UnleaRned Knowledge), a novel framework that probes for hidden retained knowledge in unlearned LLMs through adversarial suffix prompting. LURK automatically generates adversarial prompt su…

Diagnostic