arxiv.org favicon

[2403.05530] Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

5
公开标注数
5
参与人数
2026-08-12 10:14:25
首次 Whisper

本页的公开 Whisper

划选高亮2026-08-12 13:35:25
原文高亮摘录
capable of recalling and reasoning over fine-grained information
Whisper 随想笔记
Recalling fine-grained stuff is cool, but I'd rather see it not hallucinate the details first.
划选高亮2026-08-12 13:26:25
原文高亮摘录
capable of recalling and reasoning over fine-grained information
Whisper 随想笔记
10M tokens is wild, but can it actually find that one specific line in a 500-page PDF?
划选高亮2026-08-12 10:32:25
原文高亮摘录
next generation of highly compute-efficient multimodal models
Whisper 随想笔记
Still waiting for the open source model that does this without breaking the bank.
划选高亮2026-08-12 10:23:25
原文高亮摘录
next generation of highly compute-efficient multimodal models
Whisper 随想笔记
Million token context is wild, finally can throw whole codebases at it.
划选高亮2026-08-12 10:14:25
原文高亮摘录
next generation of highly compute-efficient multimodal models
Whisper 随想笔记
Compute-efficient but what about the energy cost for training these things?

分享本页 Whisper

分享到 X
短链接
https://domwhisper.com/s/79a8d5beefad
嵌入代码
<iframe src="https://domwhisper.com/embed/79a8d5beefad" width="100%" height="480" style="border:0;border-radius:16px" loading="lazy"></iframe>

看看大家在 arxiv.org 上讨论了什么

安装 DomWhisper,浏览网页时实时查看 whisper,也可以加入讨论。

获取插件