2 papers
cs.CV2025
Line of Sight: On Linear Representations in VLLMs
Achyuta Rajaram, Sarah Schwettmann, Jacob Andreas +1
Language models can be equipped with multimodal capabilities by fine-tuning on embeddings of visual inputs. But how do such multimodal models represent images in their hidden activ…
cs.CV2022
FREDSR: Fourier Residual Efficient Diffusive GAN for Single Image Super Resolution
Kyoungwan Woo, Achyuta Rajaram
FREDSR is a GAN variant that aims to outperform traditional GAN models in specific tasks such as Single Image Super Resolution with extreme parameter efficiency at the cost of per-…