[2403.05530] Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
5
公开标注数
5
参与人数
2026-08-12 10:14:25
首次 Whisper
本页的公开 Whisper
划选高亮2026-08-12 13:35:25
原文高亮摘录
“capable of recalling and reasoning over fine-grained information”
Whisper 随想笔记
Recalling fine-grained stuff is cool, but I'd rather see it not hallucinate the details first.
划选高亮2026-08-12 13:26:25
原文高亮摘录
“capable of recalling and reasoning over fine-grained information”
Whisper 随想笔记
10M tokens is wild, but can it actually find that one specific line in a 500-page PDF?
划选高亮2026-08-12 10:32:25
原文高亮摘录
“next generation of highly compute-efficient multimodal models”
Whisper 随想笔记
Still waiting for the open source model that does this without breaking the bank.
划选高亮2026-08-12 10:23:25
原文高亮摘录
“next generation of highly compute-efficient multimodal models”
Whisper 随想笔记
Million token context is wild, finally can throw whole codebases at it.
划选高亮2026-08-12 10:14:25
原文高亮摘录
“next generation of highly compute-efficient multimodal models”
Whisper 随想笔记
Compute-efficient but what about the energy cost for training these things?
分享本页 Whisper
短链接
https://domwhisper.com/s/79a8d5beefad嵌入代码
<iframe src="https://domwhisper.com/embed/79a8d5beefad" width="100%" height="480" style="border:0;border-radius:16px" loading="lazy"></iframe>