paper-with-me

홈 › Papers

DANTE: Deep AlterNations for Training nEural networks

2019-02-01 · Vaibhav B Sinha, Sneha Kudugunta, Adepu Ravi Sankar, Surya Teja Chavali, Purushottam Kar, Vineeth N. Balasubramanian

We present DANTE, a novel method for training neural networks using the alternating minimization principle. DANTE provides an alternate perspective to traditional gradient-based backpropagation techniques commonly used to train deep networks. It utilizes an adaptation of quasi-convexity to cast training a neural network as a bi-quasi-convex optimization problem. We show that for neural network configurations with both differentiable (e.g. sigmoid) and non-differentiable (e.g. ReLU) activation functions, we can perform the alternations effectively in this formulation. DANTE can also be extended to networks with multiple hidden layers. In experiments on standard datasets, neural networks trained using the proposed method were found to be promising and competitive to traditional backpropagation techniques, both in terms of quality of the solution, as well as training speed.

📄 PDF Abstract BibTeX arXiv:1902.00491

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Training Autoencoders by Alternating Minimization

2018-01-01 · ICLR 2018 1 · Sneha Kudugunta, Adepu Shankar, Surya Chavali, Vineeth Balasubramanian 외

We present DANTE, a novel method for training neural networks, in particular autoencoders, using the alternating minimization principle. DANTE provides a distinct perspective in lieu of traditional gradient-based backpro…

DANTE: A framework for mining and monitoring darknet traffic

2020-03-05 · Dvir Cohen, Yisroel Mirsky, Yuval Elovici, Rami Puzis 외

Trillions of network packets are sent over the Internet to destinations which do not exist. This 'darknet' traffic captures the activity of botnets and other malicious campaigns aiming to discover and compromise devices …

Time SeriesTime Series Analysis

A Treebank-based Approach to the Supprema Constructio in Dante’s Latin Works

2022-06-01 · LT4HALA (LREC) 2022 6 · Flavio Massimiliano Cecchini, Giulia Pedonese

This paper aims to apply a corpus-driven approach to Dante Alighieri’s Latin works using UDante, a treebank based on Dante Search and part of the Universal Dependencies project. We present a method based on the notion of…

Sentence

Desbordante: from benchmarking suite to high-performance science-intensive data profiler (preprint)

2023-01-14 · George Chernishev, Michael Polyntsov, Anton Chizhov, Kirill Stupakov 외

Pioneering data profiling systems such as Metanome and OpenClean brought public attention to science-intensive data profiling. This type of profiling aims to extract complex patterns (primitives) such as functional depen…

Benchmarking

DANTE-AD: Dual-Vision Attention Network for Long-Term Audio Description

2025-03-31 · Adrienne Deganutti, Simon Hadfield, Andrew Gilbert

Audio Description is a narrated commentary designed to aid vision-impaired audiences in perceiving key visual elements in a video. While short-form video understanding has advanced rapidly, a solution for maintaining coh…

Video DescriptionVideo UnderstandingVisual Storytelling