3 papers
cs.CL2026
The Percept-V Challenge: Can Multimodal LLMs Crack Simple Perception Problems?
Samrajnee Ghosh, Naman Agarwal, Hemanshu Garg +3
Cognitive science research treats visual perception, the ability to understand and make sense of a visual input, as one of the early developmental signs of intelligence. Its TVPS-4…
cs.SE2025
CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking
Neeva Oza, Ishaan Govil, Parul Gupta +3
LLMs have been extensively used for the task of automated code generation. In this work, we examine the applicability of LLMs for the related but relatively unexplored task of code…
cs.CV2025
GraPE: A Generate-Plan-Edit Framework for Compositional T2I Synthesis
Ashish Goswami, Satyam Kumar Modi, Santhosh Rishi Deshineni +3
Text-to-image (T2I) generation has seen significant progress with diffusion models, enabling generation of photo-realistic images from text prompts. Despite this progress, existing…