2 citations · 2 across the 2 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Large Language Models Implicitly Learn to See and Hear Just By Reading
Prateek Verma, Mert Pilanci
This paper presents a fascinating find: By training an auto-regressive LLM model on text tokens, the text model inherently develops internally an ability to understand images and a…
cs.CL2024
Adaptive Large Language Models By Layerwise Attention Shortcuts
Prateek Verma, Mert Pilanci
Transformer architectures are the backbone of the modern AI revolution. However, they are based on simply stacking the same blocks in dozens of layers and processing information se…
cs.CL2024
Towards Signal Processing In Large Language Models
Prateek Verma, Mert Pilanci
This paper introduces the idea of applying signal processing inside a Large Language Model (LLM). With the recent explosion of generative AI, our work can help bridge two fields to…