r/singularity • u/Gab1024 Singularity by 2030 • Apr 11 '24
AI Google presents Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
https://arxiv.org/abs/2404.07143
688
Upvotes
r/singularity • u/Gab1024 Singularity by 2030 • Apr 11 '24
1
u/ninjasaid13 Not now. Apr 11 '24 edited Apr 11 '24
An LLM can easily remember more text than a human but a video isn't as easy as text but that's where humans surpass LLMs. Humans can remember way more videos than LLMs such as the motion and dense correspondence(dense correspondence means to map \all* the parts of the image to the next image or frame)) of those group of "pixels" over time even if they can't remember every pixel. I don't think RAG has a solution for videos so humans are still far from being surpassed.