paper-with-me

Papers

Data Dependent Kernel Approximation using Pseudo Random Fourier Features

2017-11-27 · Bharath Bhushan Damodaran, Nicolas Courty, Philippe-Henri Gosselin

Kernel methods are powerful and flexible approach to solve many problems in machine learning. Due to the pairwise evaluations in kernel methods, the complexity of kernel computation grows as the data size increases; thus the applicability of kernel methods is limited for large scale datasets. Random Fourier Features (RFF) has been proposed to scale the kernel method for solving large scale datasets by approximating kernel function using randomized Fourier features. While this method proved very popular, still it exists shortcomings to be effectively used. As RFF samples the randomized features from a distribution independent of training data, it requires sufficient large number of feature expansions to have similar performances to kernelized classifiers, and this is proportional to the number samples in the dataset. Thus, reducing the number of feature dimensions is necessary to effectively scale to large datasets. In this paper, we propose a kernel approximation method in a data dependent way, coined as Pseudo Random Fourier Features (PRFF) for reducing the number of feature dimensions and also to improve the prediction performance. The proposed approach is evaluated on classification and regression problems and compared with the RFF, orthogonal random features and Nystr{\"o}m approach

📄 PDF Abstract BibTeX arXiv:1711.09783

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Randomized Independent Component Analysis

2016-09-22 · Matan Sela, Ron Kimmel

Independent component analysis (ICA) is a method for recovering statistically independent signals from observations of unknown linear combinations of the sources. Some of the most accurate ICA decomposition methods requi…

Data-dependent compression of random features for large-scale kernel approximation

2018-10-09 · Raj Agrawal, Trevor Campbell, Jonathan H. Huggins, Tamara Broderick

Kernel methods offer the flexibility to learn complex relationships in modern, large data sets while enjoying strong theoretical guarantees on quality. Unfortunately, these methods typically require cubic running time in…

feature selection

The Mondrian Kernel

2016-06-16 · Matej Balog, Balaji Lakshminarayanan, Zoubin Ghahramani, Daniel M. Roy 외

We introduce the Mondrian kernel, a fast random feature approximation to the Laplace kernel. It is suitable for both batch and online learning, and admits a fast kernel-width-selection procedure as the random features ca…

Approximate Kernel PCA Using Random Features: Computational vs. Statistical Trade-off

2017-06-20 · Bharath Sriperumbudur, Nicholas Sterge

Kernel methods are powerful learning methodologies that allow to perform non-linear data analysis. Despite their popularity, they suffer from poor scalability in big data scenarios. Various approximation methods, includi…

ORCCA: Optimal Randomized Canonical Correlation Analysis

2019-10-11 · Yinsong Wang, Shahin Shahrampour

Random features approach has been widely used for kernel approximation in large-scale machine learning. A number of recent studies have explored data-dependent sampling of features, modifying the stochastic oracle from w…

BIG-bench Machine Learningscoring rule