Comum · Saberes
Batch normalization
Normalization technique used to make training faster and more stable by adjusting the inputs to each layer, recentering them around zero and rescaling them to a standard size
In artificial neural networks, batch normalization (also known as batch norm) is a normalization technique used to make training faster and more stable by adjusting the inputs to each layer—re-centering them around zero and re-scaling them to a standard size. It was introduced by Sergey Ioffe and Christian Szegedy in 2015.
Na Wikipédia
Texto em inglês Ainda não há artigo no seu idioma: trecho em inglês.
In artificial neural networks, batch normalization (also known as batch norm) is a normalization technique used to make training faster and more stable by adjusting the inputs to each layer—re-centering them around zero and re-scaling them to a standard size. It was introduced by Sergey Ioffe and Christian Szegedy in 2015. Experts still debate why batch normalization works so well. It was initially thought to tackle internal covariate shift, a problem where parameter initialization and changes in the distribution of the inputs of each layer affect the learning rate of the network. However, newer research suggests it does not fix this shift but instead smooths the objective function—a mathematical guide the network follows to improve—enhancing performance. In very deep networks, batch normalization can initially cause a severe gradient explosion—where updates to the network grow uncontrollably large—but this is managed with shortcuts called skip connections in residual networks. Another theory is that batch normalization adjusts data by handling its size and path separately, speeding up training.
Texto: Wikipédia em inglês, CC BY-SA 4.0. ·
Cartas próximas
-
V★
Validated numerics
Numerics including mathematically strict error evaluation
-
S★
Siamese neural network
Form of neural network
-
T★
Teoria da Inferência Indutiva de Solomonoff
-
U★★
U-Net
Rede neural convolucional desenvolvida na Universidade de Friburgo
-
★
Jackknife resampling
Statistical method for resampling
-
A★
AdaBoost
-
★
agregação bootstrap
Algoritmo de aprendizagem de máquina
-
★
Leaky Bucket
-
A★★
Algoritmo de Markov
-
L★
Learning with errors
Problem in machine learning that is conjectured to be hard to solve. Introduced by Oded Regev in 2005, it is a generalization of the parity learning problem
-
E★★★
Espaço de Banach
-
★★★
Método de Newton–Raphson
Algoritmo para encontrar raízes
-
★★★★
Transformer (aprendizado profundo)
Um modelo de aprendizagem de máquina do Google Brain
-
C★★
Codificação preditiva
-
S★★
Standard part function
In non-standard analysis, the standard part function is a function from the limited (finite) hyperreal numbers to the real numbers.
-
N★★
Número normal
Número real cujos algarismos são distribuídos de maneira aleatória no seu desenvolvimento em qualquer base
-
C★
Change of variables
Technique in algebra in which the original variables are replaced with functions of other variables
-
B★
Basic Linear Algebra Subprograms
Routines for performing common linear algebra operations