paper-with-me

홈 › Papers

Improving neural networks with bunches of neurons modeled by Kumaraswamy units: Preliminary study

2015-05-11 · Jakub Mikolaj Tomczak

Deep neural networks have recently achieved state-of-the-art results in many machine learning problems, e.g., speech recognition or object recognition. Hitherto, work on rectified linear units (ReLU) provides empirical and theoretical evidence on performance increase of neural networks comparing to typically used sigmoid activation function. In this paper, we investigate a new manner of improving neural networks by introducing a bunch of copies of the same neuron modeled by the generalized Kumaraswamy distribution. As a result, we propose novel non-linear activation function which we refer to as Kumaraswamy unit which is closely related to ReLU. In the experimental study with MNIST image corpora we evaluate the Kumaraswamy unit applied to single-layer (shallow) neural network and report a significant drop in test classification error and test cross-entropy in comparison to sigmoid unit, ReLU and Noisy ReLU.

📄 PDF Abstract BibTeX arXiv:1505.02581

Code (0)

등록된 구현이 없습니다.

Tasks

Object Recognitionspeech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Sigmoid Activation 설명 없음

Similar Papers 제목 키워드 기반

Automatic Detection, Positioning and Counting of Grape Bunches Using Robots

2024-12-12 · Xumin Gao

In order to promote agricultural automatic picking and yield estimation technology, this project designs a set of automatic detection, positioning and counting algorithms for grape bunches, and applies it to agricultural…

Position

The Knowledge Microscope: Features as Better Analytical Lenses than Neurons

2025-02-18 · YuHeng Chen, Pengfei Cao, Kang Liu, Jun Zhao

Previous studies primarily utilize MLP neurons as units of analysis for understanding the mechanisms of factual knowledge in Language Models (LMs); however, neurons suffer from polysemanticity, leading to limited knowled…

Two-argument activation functions learn soft XOR operations like cortical neurons

2021-10-13 · KiJung Yoon, Emin Orhan, Juhyun Kim, Xaq Pitkow

Neurons in the brain are complex machines with distinct functional compartments that interact nonlinearly. In contrast, neurons in artificial neural networks abstract away this complexity, typically down to a scalar acti…

Vocal Bursts Valence Prediction

Stabilizing the Kumaraswamy Distribution

2024-10-01 · Max Wasserman, Gonzalo Mateos

Large-scale latent variable models require expressive continuous distributions that support efficient sampling and low-variance differentiation, achievable through the reparameterization trick. The Kumaraswamy (KS) distr…

Link PredictionMulti-Armed BanditsUncertainty Quantification

Binding threshold units with artificial oscillatory neurons

2025-05-06 · Vladimir Fanaskov, Ivan Oseledets

Artificial Kuramoto oscillatory neurons were recently introduced as an alternative to threshold units. Empirical evidence suggests that oscillatory units outperform threshold units in several tasks including unsupervised…

FormObject Discovery