Batch normalization
Normalization technique used to make training faster and more stable by adjusting the inputs to each layer, recentering them around zero and rescaling them to a standard size
In artificial neural networks, batch normalization (also known as batch norm) is a normalization technique used to make training faster and more stable by adjusting the inputs to each layer—re-centering them around zero and re-scaling them to a standard size. It was introduced by Sergey Ioffe and Christian Szegedy in 2015.
Nº Q55080248 ★
Common · Knowledge
Batch normalization
Normalization technique used to make training faster and more stable by adjusting the inputs to each layer, recentering them around zero and rescaling them to a standard size
In artificial neural networks, batch normalization (also known as batch norm) is a normalization technique used to make training faster and more stable by adjusting the inputs to each layer—re-centering them around zero and re-scaling them to a standard size. It was introduced by Sergey Ioffe and Christian Szegedy in 2015.
From Wikipedia
In artificial neural networks, batch normalization (also known as batch norm) is a normalization technique used to make training faster and more stable by adjusting the inputs to each layer—re-centering them around zero and re-scaling them to a standard size. It was introduced by Sergey Ioffe and Christian Szegedy in 2015. Experts still debate why batch normalization works so well. It was initially thought to tackle internal covariate shift, a problem where parameter initialization and changes in the distribution of the inputs of each layer affect the learning rate of the network. However, newer research suggests it does not fix this shift but instead smooths the objective function—a mathematical guide the network follows to improve—enhancing performance. In very deep networks, batch normalization can initially cause a severe gradient explosion—where updates to the network grow uncontrollably large—but this is managed with shortcuts called skip connections in residual networks. Another theory is that batch normalization adjusts data by handling its size and path separately, speeding up training.
Text: Wikipédia, CC BY-SA 4.0. ·
Related cards
-
F
Fourth normal form
Normal form used in database normalization concerned with multivalued dependency
Nº Q2492261 ★
Not listed
-
Connectionism
Approach in cognitive science that hopes to explain mental phenomena using artificial neural networks
Nº Q203790 ★
Not listed
-
Box–Muller transform
Statistical transform
Nº Q895514 ★
Not listed
-
I
Intelligence amplification
Augmentation of intelligence through the use of information technology
Nº Q1362688 ★★
Not listed
-
E
Efficiently updatable neural network
Neural network-based evaluation function
Nº Q98078848 ★
Not listed
-
L
Latent diffusion model
Deep generative model
Nº Q130641540 ★
Not listed
-
S
Smith normal form
Normal form for a matrix with values in a principal ideal domain
Nº Q7545384 ★
Not listed
-
N
Normalization (statistics)
Statistical procedure
Nº Q249772 ★★
Not listed
-
Stochastic resonance
Signal boosting phenomenon using white noise
Nº Q1999781 ★
Not listed
-
J
Job scheduler
Computer application for controlling unattended background program execution of jobs
Nº Q1641413 ★
Not listed
-
TensorFlow
Machine learning software framework
Nº Q21447895 ★★★
Not listed
-
Kalman filter
Algorithm that estimates unknowns from a series of measurements over time
Nº Q846780 ★★★★
Not listed
-
B
BIRCH
Clustering algorithm
Nº Q4835721 ★★
Not listed
-
H
Horner's method
Algorithm for polynomial evaluation
Nº Q944658 ★★
Not listed
-
L
Linear-feedback shift register
Type of shift register in computing
Nº Q681101 ★★
Not listed
-
MAFFT
Multiple alignment software for amino acid or nucleotide sequences
Nº Q6714151 ★
Not listed
-
L
Loop unrolling
Loop transformation technique
Nº Q1869750 ★
Not listed
-
B
Backward Euler method
Numerical method for solving differential equations
Nº Q2736820 ★★
Not listed