paper-with-me

Papers

Organizing Unstructured Image Collections using Natural Language

2024-10-07 · Mingxuan Liu, Zhun Zhong, Jun Li, Gianni Franchi, Subhankar Roy, Elisa Ricci

Organizing unstructured visual data into semantic clusters is a key challenge in computer vision. Traditional deep clustering (DC) approaches focus on a single partition of data, while multiple clustering (MC) methods address this limitation by uncovering distinct clustering solutions. The rise of large language models (LLMs) and multimodal LLMs (MLLMs) has enhanced MC by allowing users to define clustering criteria in natural language. However, manually specifying criteria for large datasets is impractical. In this work, we introduce the task Semantic Multiple Clustering (SMC) that aims to automatically discover clustering criteria from large image collections, uncovering interpretable substructures without requiring human input. Our framework, Text Driven Semantic Multiple Clustering (TeDeSC), uses text as a proxy to concurrently reason over large image collections, discover partitioning criteria, expressed in natural language, and reveal semantic substructures. To evaluate TeDeSC, we introduce the COCO-4c and Food-4c benchmarks, each containing four grouping criteria and ground-truth annotations. We apply TeDeSC to various applications, such as discovering biases and analyzing social media image popularity, demonstrating its utility as a tool for automatically organizing image collections and revealing novel insights.

📄 PDF Abstract BibTeX arXiv:2410.05217

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringDeep Clustering

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

T2K\textasciicircum2: a System for Automatically Extracting and Organizing Knowledge from Texts

2014-05-01 · LREC 2014 5 · Felice Dell{'}Orletta, Giulia Venturi, Andrea Cimino, Simonetta Montemagni

In this paper, we present T2K{\textasciicircum}2, a suite of tools for automatically extracting domain―specific knowledge from collections of Italian and English texts. T2K{\textasciicircum}2 (Text―To―Knowledge v2) relie…

SR-Clustering: Semantic Regularized Clustering for Egocentric Photo Streams Segmentation

2015-12-22 · Mariella Dimiccoli, Marc Bolaños, Estefania Talavera, Maedeh Aghaei 외

While wearable cameras are becoming increasingly popular, locating relevant information in large unstructured collections of egocentric images is still a tedious and time consuming processes. This paper addresses the pro…

Clustering

UQE: A Query Engine for Unstructured Databases

2024-06-23 · Hanjun Dai, Bethany Yixin Wang, Xingchen Wan, Bo Dai 외

Analytics on structured data is a mature field with many successful methods. However, most real world data exists in unstructured form, such as images and conversations. We investigate the potential of Large Language Mod…

Semantic Retrieval

Behavior Discovery and Alignment of Articulated Object Classes from Unstructured Video

2015-11-30 · Luca Del Pero, Susanna Ricco, Rahul Sukthankar, Vittorio Ferrari

We propose an automatic system for organizing the content of a collection of unstructured videos of an articulated object class (e.g. tiger, horse). By exploiting the recurring motion patterns of the class across videos,…

Retrieval

TopicTag: Automatic Annotation of NMF Topic Models Using Chain of Thought and Prompt Tuning with LLMs

2024-07-29 · Selma Wanna, Ryan Barron, Nick Solovyev, Maksim E. Eren 외

Topic modeling is a technique for organizing and extracting themes from large collections of unstructured text. Non-negative matrix factorization (NMF) is a common unsupervised approach that decomposes a term frequency-i…

Knowledge GraphsManagementPrompt EngineeringTopic Models