Simple and Effective Stochastic Neural Networks

Tianyuan Yu; Yongxin Yang; Da Li; Timothy Hospedales; Tao Xiang

doi:10.1609/aaai.v35i4.16436

Authors

Tianyuan Yu University of Surrey
Yongxin Yang University of Surrey
Da Li University of Edinburgh Samsung AI Centre
Timothy Hospedales University of Edinburgh Samsung AI Centre
Tao Xiang University of Surrey

DOI:

https://doi.org/10.1609/aaai.v35i4.16436

Keywords:

Applications

Abstract

Stochastic neural networks (SNNs) are currently topical, with several paradigms being actively investigated including dropout, Bayesian neural networks, variational information bottleneck (VIB) and noise regularized learning. These neural network variants impact several major considerations, including generalization, network compression, robustness against adversarial attack and label noise, and model calibration. However, many existing networks are complicated and expensive to train, and/or only address one or two of these practical considerations. In this paper we propose a simple and effective stochastic neural network (SE-SNN) architecture for discriminative learning by directly modeling activation uncertainty and encouraging high activation variability. Compared to existing SNNs, our SE-SNN is simpler to implement and faster to train, and produces state of the art results on network compression by pruning, adversarial defense, learning with label noise, and model calibration.

Simple and Effective Stochastic Neural Networks

Authors

DOI:

Keywords:

Abstract

Downloads

Published

How to Cite

Issue

Section

Information

Developed By

Subscription