arxiv.org favicon

[2403.05530] Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

[2403.05530] Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

#2
273
Public whispers
7
Contributors
2026-06-28 09:09:18
First whispered

Discussion activity

Public whispers on arxiv.org over the last 17 weeks

26 active days
Less
More

Recent public whispers

RSS
Text Highlight2026-08-12 13:35:25
Original Highlight Excerpt
"capable of recalling and reasoning over fine-grained information"
Whisper Note
Recalling fine-grained stuff is cool, but I'd rather see it not hallucinate the details first.
Text Highlight2026-08-12 13:26:25
Original Highlight Excerpt
"capable of recalling and reasoning over fine-grained information"
Whisper Note
10M tokens is wild, but can it actually find that one specific line in a 500-page PDF?
Text Highlight2026-08-12 10:32:25
Original Highlight Excerpt
"next generation of highly compute-efficient multimodal models"
Whisper Note
Still waiting for the open source model that does this without breaking the bank.
Text Highlight2026-08-12 10:23:25
Original Highlight Excerpt
"next generation of highly compute-efficient multimodal models"
Whisper Note
Million token context is wild, finally can throw whole codebases at it.
Text Highlight2026-08-12 10:14:25
Original Highlight Excerpt
"next generation of highly compute-efficient multimodal models"
Whisper Note
Compute-efficient but what about the energy cost for training these things?
Text Highlight2026-08-11 18:38:31
Original Highlight Excerpt
"QLoRA backprop"
Whisper Note
My 48GB card finally feels useful for something other than gaming.
Text Highlight2026-08-11 03:06:31
Original Highlight Excerpt
"General reasoning represents a long-standing and formidable challenge"
Whisper Note
The abstract is truncated, but DeepSeek-R1's RL approach sounds really promising.
Text Highlight2026-08-10 14:32:38
Original Highlight Excerpt
"a herd of language models that natively support multilingua"
Whisper Note
multilinguality is cool but tool usage is where it's really at
Text Highlight2026-08-10 14:23:38
Original Highlight Excerpt
"a herd of language models that natively support multilingua"
Whisper Note
405B params is insane, wonder how much compute that took
Text Highlight2026-08-10 14:14:38
Original Highlight Excerpt
"a herd of language models that natively support multilingua"
Whisper Note
finally a model that actually gets multiple languages without breaking a sweat
Text Highlight2026-07-28 16:29:34
Original Highlight Excerpt
"reduce the training time"
Whisper Note
Finally a normalization that works for RNNs without all that batch size fuss.
Text Highlight2026-07-28 16:07:22
Original Highlight Excerpt
"arXivLabs: experimental projects with community collaborators"
Whisper Note
Is this just for the staff or can anyone pitch in?
Text Highlight2026-07-28 16:05:55
Original Highlight Excerpt
"arXivLabs: experimental projects with community collaborators"
Whisper Note
Cool, community collabs always bring fresh ideas to the table.
Text Highlight2026-07-28 13:26:34
Original Highlight Excerpt
"batch norma"
Whisper Note
Does this actually scale well for CNNs though? Batch norm felt more natural there.
Text Highlight2026-07-28 13:17:34
Original Highlight Excerpt
"batch norma"
Whisper Note
So layer norm is basically batch norm but for single examples, makes sense for RNNs.
Text Highlight2026-07-28 13:04:22
Original Highlight Excerpt
"use of consistency training on a large amount of unlabe"
Whisper Note
So basically, they found that how you add noise matters more than the amount of unlabeled data?
Text Highlight2026-07-28 13:02:55
Original Highlight Excerpt
"arXiv-issued DOI via DataCite"
Whisper Note
Huh, I always thought DOIs were just for journals, cool that arXiv does it too.
Text Highlight2026-07-28 12:55:22
Original Highlight Excerpt
"use of consistency training on a large amount of unlabe"
Whisper Note
Consistency training with better noise is the real winner here, not just the semi-supervised part.
Text Highlight2026-07-28 12:53:55
Original Highlight Excerpt
"arXiv-issued DOI via DataCite"
Whisper Note
Wait, arXiv is giving DOIs now? That's actually kinda neat.
Text Highlight2026-07-28 10:23:34
Original Highlight Excerpt
"normalize the activities of the neurons"
Whisper Note
tried it on my lstm and yeah, training felt way smoother

See what people are saying on arxiv.org

Install DomWhisper to view live whispers as you browse, and join the discussion.

Get the extension