Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
CS-VLM: Compressed Sensing Attention for Efficient Vision-Language Representation Learning
Andrew Kiruluta, Preethi Raju, Priscilla Burity
Vision-Language Models (vLLMs) have emerged as powerful architectures for joint reasoning over visual and textual inputs, enabling breakthroughs in image captioning, cross modal re…
cs.CV2025
From Pixels and Words to Waves: A Unified Framework for Spectral Dictionary vLLMs
Andrew Kiruluta, Priscilla Burity
Vision-language models (VLMs) unify computer vision and natural language processing in a single architecture capable of interpreting and describing images. Most state-of-the-art sy…