arxiv.org favicon

[1502.03167] Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift

5
Public whispers
5
Contributors
2026-07-27 10:05:30
First whispered

Public whispers on this page

Text Highlight2026-07-27 13:26:30
Original Highlight Excerpt
"Reducing Internal Covariate Shift"
Whisper Note
Still not sure if covariate shift is the real reason it works, but it does.
Text Highlight2026-07-27 13:17:30
Original Highlight Excerpt
"Reducing Internal Covariate Shift"
Whisper Note
This paper basically changed how we train deep nets overnight.
Text Highlight2026-07-27 10:23:30
Original Highlight Excerpt
"distribution of each layer's inputs changes during training"
Whisper Note
I remember when this came out, total game changer for my CNN experiments.
Text Highlight2026-07-27 10:14:30
Original Highlight Excerpt
"distribution of each layer's inputs changes during training"
Whisper Note
Skeptical it works that well without careful tuning, but results speak.
Text Highlight2026-07-27 10:05:30
Original Highlight Excerpt
"distribution of each layer's inputs changes during training"
Whisper Note
Honestly this is the paper that made deep nets actually trainable for me.

Share this page's whispers

Share to X
Short link
https://domwhisper.com/s/df9c222d88c4
Embed snippet
<iframe src="https://domwhisper.com/embed/df9c222d88c4" width="100%" height="480" style="border:0;border-radius:16px" loading="lazy"></iframe>

See what people are discussing on arxiv.org

Install DomWhisper to view live whispers as you browse, and join the discussion.

Get the extension