A Feature Selection Method for Multivariate Performance Measures
Feature selection with specific multivariate performance measures is the key to the success of many applications, such as image retrieval and text classification. The existing feature selection methods are usually designed for classification error. In this paper, we propose a generalized sparse regularizer. Based on the proposed regularizer, we present a unified feature selection framework for general loss functions. In particular, we study the novel feature selection paradigm by optimizing multivariate performance measures. The resultant formulation is a challenging problem for high-dimensional data. Hence, a two-layer cutting plane algorithm is proposed to solve this problem, and the convergence is presented. In addition, we adapt the proposed method to optimize multivariate measures for multiple instance learning problems. The analyses by comparing with the state-of-the-art feature selection methods show that the proposed method is superior to others. Extensive experiments on large-scale and high-dimensional real world datasets show that the proposed method outperforms $l_1$-SVM and SVM-RFE when choosing a small subset of features, and achieves significantly improved performances over SVM$^{perf}$ in terms of $F_1$-score.
Code (0)
등록된 구현이 없습니다.
Tasks
feature selectionGeneral ClassificationImage RetrievalMultiple Instance LearningRetrievaltext-classificationText ClassificationSimilar Papers 제목 키워드 기반
micompm: A MATLAB/Octave toolbox for multivariate independent comparison of observations
micompm is a MATLAB / GNU Octave port of the original micompr R package for comparing multivariate samples associated with different groups. Its purpose is to determine if the compared samples are significantly different…
Time SeriesTime Series AnalysisMultivariate risk measures: a constructive approach based on selections
Since risky positions in multivariate portfolios can be offset by various choices of capital requirements that depend on the exchange rules and related transaction costs, it is natural to assume that the risk measures of…
Markov Blanket Ranking using Kernel-based Conditional Dependence Measures
Developing feature selection algorithms that move beyond a pure correlational to a more causal analysis of observational data is an important problem in the sciences. Several algorithms attempt to do so by discovering th…
feature selectionMulti-view learning for multivariate performance measures optimization
In this paper, we propose the problem of optimizing multivariate performance measures from multi-view data, and an effective method to solve it. This problem has two features: the data points are presented by multiple vi…
MULTI-VIEW LEARNINGGlobal Sensitivity Analysis with Dependence Measures
Global sensitivity analysis with variance-based measures suffers from several theoretical and practical limitations, since they focus only on the variance of the output and handle multivariate variables in a limited way.…
feature selectionSensitivity