paper-with-me

홈 › Papers

AMITE: A Novel Polynomial Expansion for Analyzing Neural Network Nonlinearities

2020-07-13 · Mauro J. Sanchirico III, Xun Jiao, C. Nataraj

Polynomial expansions are important in the analysis of neural network nonlinearities. They have been applied thereto addressing well-known difficulties in verification, explainability, and security. Existing approaches span classical Taylor and Chebyshev methods, asymptotics, and many numerical approaches. We find that while these individually have useful properties such as exact error formulas, adjustable domain, and robustness to undefined derivatives, there are no approaches that provide a consistent method yielding an expansion with all these properties. To address this, we develop an analytically modified integral transform expansion (AMITE), a novel expansion via integral transforms modified using derived criteria for convergence. We show the general expansion and then demonstrate application for two popular activation functions, hyperbolic tangent and rectified linear units. Compared with existing expansions (i.e., Chebyshev, Taylor, and numerical) employed to this end, AMITE is the first to provide six previously mutually exclusive desired expansion properties such as exact formulas for the coefficients and exact expansion errors (Table II). We demonstrate the effectiveness of AMITE in two case studies. First, a multivariate polynomial form is efficiently extracted from a single hidden layer black-box Multi-Layer Perceptron (MLP) to facilitate equivalence testing from noisy stimulus-response pairs. Second, a variety of Feed-Forward Neural Network (FFNN) architectures having between 3 and 7 layers are range bounded using Taylor models improved by the AMITE polynomials and error formulas. AMITE presents a new dimension of expansion methods suitable for analysis/approximation of nonlinearities in neural networks, opening new directions and opportunities for the theoretical analysis and systematic testing of neural networks.

📄 PDF Abstract BibTeX arXiv:2007.06226

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Chebyshev-Augmented One-Shot Transfer Learning for PINNs on Nonlinear Differential Equations

2026-05-02 · Yiqi Rao, Pavlos Protopapas arxiv

Physics-Informed Neural Networks (PINNs) offer a flexible paradigm for solving differential equations by embedding governing laws into the training objective. A persistent limitation is instance specificity: standard PIN…

Transfer Learning

Multielement polynomial chaos Kriging-based metamodelling for Bayesian inference of non-smooth systems

2022-12-05 · J. C. García-Merino, C. Calvo-Jurado, E. Martínez-Pañeda, E. García-Macías

This paper presents a surrogate modelling technique based on domain partitioning for Bayesian parameter inference of highly nonlinear engineering models. In order to alleviate the computational burden typically involved …

Bayesian Inference

DNAMite: Interpretable Calibrated Survival Analysis with Discretized Additive Models

2024-11-08 · Mike Van Ness, Billy Block, Madeleine Udell

Survival analysis is a classic problem in statistics with important applications in healthcare. Most machine learning models for survival analysis are black-box models, limiting their use in healthcare settings where int…

Additive modelsSurvival Analysis

DynaMITE-RL: A Dynamic Model for Improved Temporal Meta-Reinforcement Learning

2024-02-25 · Anthony Liang, Guy Tennenholtz, Chih-Wei Hsu, Yinlam Chow 외

We introduce DynaMITE-RL, a meta-reinforcement learning (meta-RL) approach to approximate inference in environments where the latent state evolves at varying rates. We model episode sessions - parts of the episode where …

continuous-controlContinuous ControlMeta Reinforcement Learning

Sign Clustering and Topic Extraction in Proto-Elamite

2019-06-01 · WS 2019 6 · Logan Born, Kate Kelley, Nishant Kambhatla, Carolyn Chen 외

We describe a first attempt at using techniques from computational linguistics to analyze the undeciphered proto-Elamite script. Using hierarchical clustering, n-gram frequencies, and LDA topic models, we both replicate …

ClusteringDeciphermentTopic Models