paper-with-me

Papers

Sparse Output Coding for Large-Scale Visual Recognition

2013-06-01 · CVPR 2013 6 · Bin Zhao, Eric P. Xing

Many vision tasks require a multi-class classifier to discriminate multiple categories, on the order of hundreds or thousands. In this paper, we propose sparse output coding, a principled way for large-scale multi-class classification, by turning high-cardinality multi-class categorization into a bit-by-bit decoding problem. Specifically, sparse output coding is composed of two steps: efficient coding matrix learning with scalability to thousands of classes, and probabilistic decoding. Empirical results on object recognition and scene classification demonstrate the effectiveness of our proposed approach.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral ClassificationMulti-class ClassificationObject RecognitionScene Classification

Similar Papers 제목 키워드 기반

VASparse: Towards Efficient Visual Hallucination Mitigation for Large Vision-Language Model via Visual-Aware Sparsification

2025-01-11 · Xianwei Zhuang, Zhihong Zhu, Yuxin Xie, Liming Liang 외

Large Vision-Language Models (LVLMs) may produce outputs that are unfaithful to reality, also known as visual hallucinations (VH), which significantly impedes their real-world usage. To alleviate VH, various decoding str…

HallucinationLanguage ModelingLanguage Modelling

VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsification

2025-01-01 · CVPR 2025 1 · Xianwei Zhuang, Zhihong Zhu, Yuxin Xie, Liming Liang 외

Large Vision-Language Models (LVLMs) may produce outputs that are unfaithful to reality, also known as visual hallucinations (VH), which significantly impedes their real-world usage. To alleviate VH, various decoding…

Hallucination

Generalized Lasso based Approximation of Sparse Coding for Visual Recognition

2011-12-01 · NeurIPS 2011 12 · Nobuyuki Morioka, Shin'ichi Satoh

Sparse coding, a method of explaining sensory data with as few dictionary bases as possible, has attracted much attention in computer vision. For visual object category recognition, L1 regularized sparse coding is combin…

Object Recognition

Inference-Free Multimodal Learned Sparse Retrieval for Production-Scale Visual Document Search

2026-05-29 · Gyu-Hwung Cho, Youngjune Lee, Kiyoon Jeong, Siyoung Lee 외 arxiv

As large-scale visual-document corpora such as arXiv papers and enterprise PDFs continue to grow, visual-document retrieval has gained increasing attention; yet it still lacks a deployable system that lexically indexes v…

Group Sparse Coding with a Laplacian Scale Mixture Prior

2010-12-01 · NeurIPS 2010 12 · Pierre Garrigues, Bruno A. Olshausen

We propose a class of sparse coding models that utilizes a Laplacian Scale Mixture (LSM) prior to model dependencies among coefficients. Each coefficient is modeled as a Laplacian distribution with a variable scale param…

Compressive Sensing