paper-with-me

Papers

Improving Image Clustering using Sparse Text and the Wisdom of the Crowds

2014-05-08 · Anna Ma, Arjuna Flenner, Deanna Needell, Allon G. Percus

We propose a method to improve image clustering using sparse text and the wisdom of the crowds. In particular, we present a method to fuse two different kinds of document features, image and text features, and use a common dictionary or "wisdom of the crowds" as the connection between the two different kinds of documents. With the proposed fusion matrix, we use topic modeling via non-negative matrix factorization to cluster documents.

📄 PDF Abstract BibTeX arXiv:1405.2102

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringImage Clustering

Similar Papers 제목 키워드 기반

WoCE: a framework for clustering ensemble by exploiting the wisdom of Crowds theory

2016-12-20 · Muhammad Yousefnezhad, Sheng-Jun Huang, Daoqiang Zhang

The Wisdom of Crowds (WOC), as a theory in the social science, gets a new paradigm in computer science. The WOC theory explains that the aggregate decision made by a group is often better than those of its individual mem…

ClusteringClustering EnsembleDiversity

Approximating Wisdom of Crowds using K-RBMs

2016-11-16 · Abhay Gupta

An important way to make large training sets is to gather noisy labels from crowds of non experts. We propose a method to aggregate noisy labels collected from a crowd of workers or annotators. Eliciting labels is import…

Clustering

CrowdSelect: Synthetic Instruction Data Selection with Multi-LLM Wisdom

2025-03-03 · Yisen Li, Lingfeng Yang, Wenxuan Shen, Pan Zhou 외

Distilling advanced Large Language Models' instruction-following capabilities into smaller models using a selected subset has become a mainstream approach in model training. While existing synthetic instruction data sele…

Instruction Following

Foolish Crowds Support Benign Overfitting

2021-10-06 · Niladri S. Chatterji, Philip M. Long

We prove a lower bound on the excess risk of sparse interpolating procedures for linear regression with Gaussian data in the overparameterized regime. We apply this result to obtain a lower bound for basis pursuit (the m…

regression

Wisdom from Diversity: Bias Mitigation Through Hybrid Human-LLM Crowds

2025-05-18 · Axel Abels, Tom Lenaerts

Despite their performance, large language models (LLMs) can inadvertently perpetuate biases found in the data they are trained on. By analyzing LLM responses to bias-eliciting headlines, we find that these models often m…

Diversity