arxiv.org favicon

[1607.06450] Layer Normalization

6
Public whispers
6
Contributors
2026-07-28 10:05:34
First whispered

Public whispers on this page

Text Highlight2026-07-28 16:29:34
Original Highlight Excerpt
"reduce the training time"
Whisper Note
Finally a normalization that works for RNNs without all that batch size fuss.
Text Highlight2026-07-28 13:26:34
Original Highlight Excerpt
"batch norma"
Whisper Note
Does this actually scale well for CNNs though? Batch norm felt more natural there.
Text Highlight2026-07-28 13:17:34
Original Highlight Excerpt
"batch norma"
Whisper Note
So layer norm is basically batch norm but for single examples, makes sense for RNNs.
Text Highlight2026-07-28 10:23:34
Original Highlight Excerpt
"normalize the activities of the neurons"
Whisper Note
tried it on my lstm and yeah, training felt way smoother
Text Highlight2026-07-28 10:14:34
Original Highlight Excerpt
"normalize the activities of the neurons"
Whisper Note
but does it actually help with small datasets or just big ones?
Text Highlight2026-07-28 10:05:34
Original Highlight Excerpt
"normalize the activities of the neurons"
Whisper Note
finally something that works for rnns without batch size headaches

Share this page's whispers

Share to X
Short link
https://domwhisper.com/s/c0695dd7ff75
Embed snippet
<iframe src="https://domwhisper.com/embed/c0695dd7ff75" width="100%" height="480" style="border:0;border-radius:16px" loading="lazy"></iframe>

See what people are discussing on arxiv.org

Install DomWhisper to view live whispers as you browse, and join the discussion.

Get the extension