82 citations · 123 across the 4 of their papers we have counts for
4 papers
Cosmos 3: Omnimodal World Models for Physical AI
NVIDIA, :, Aditi +293
We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences within a unified mixture-of-t…
MELODI: Exploring Memory Compression for Long Contexts
Yinpeng Chen, DeLesley Hutchins, Aren Jansen +3
We present MELODI, a novel memory architecture designed to efficiently process long documents using short context windows. The key principle behind MELODI is to represent short-ter…
Memorizing Transformers
Yuhuai Wu, Markus N. Rabe, DeLesley Hutchins +1
Language models typically need to be trained or finetuned in order to acquire new knowledge, which involves updating their weights. We instead envision language models that can sim…
Deep Learning with Dynamic Computation Graphs
Moshe Looks, Marcello Herreshoff, DeLesley Hutchins +1
Neural networks that compute over graph structures are a natural fit for problems in a variety of domains, including natural language (parse trees) and cheminformatics (molecular g…