Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Explore this paper's citation graph

Summary

Gemini 1.5 models achieve near-perfect recall on long-context retrieval tasks across modalities, improve the state-of-the-art in long-document QA, long-video QA and long-context ASR, and match or surpass Gemini 1.0 Ultra's state-of-the-art performance across a broad set of benchmarks.

Type
preprint
Published
2024-03-08
Cited by
3,933
References
0
Access
Open access

Keywords

Context (archaeology), Business, Computer science, Internet privacy, Geography

References

No references recorded for this paper.

Cited by

Related papers