2 papers
cs.CV2025
DPAR: Dynamic Patchification for Efficient Autoregressive Visual Generation
Divyansh Srivastava, Akshay Mehra, Pranav Maneriker +5
Decoder-only autoregressive image generation typically relies on fixed-length tokenization schemes whose token counts grow quadratically with resolution, substantially increasing t…
cs.CV2025
VLG-CBM: Training Concept Bottleneck Models with Vision-Language Guidance
Divyansh Srivastava, Ge Yan, Tsui-Wei Weng
Concept Bottleneck Models (CBMs) provide interpretable prediction by introducing an intermediate Concept Bottleneck Layer (CBL), which encodes human-understandable concepts to expl…