3 papers
cs.CL2026
The Percept-V Challenge: Can Multimodal LLMs Crack Simple Perception Problems?
Samrajnee Ghosh, Naman Agarwal, Hemanshu Garg +3
Cognitive science research treats visual perception, the ability to understand and make sense of a visual input, as one of the early developmental signs of intelligence. Its TVPS-4…
cs.AI2025
FCoReBench: Can Large Language Models Solve Challenging First-Order Combinatorial Reasoning Problems?
Chinmay Mittal, Krishna Kartik, Mausam +1
Can the large language models (LLMs) solve challenging first-order combinatorial reasoning problems such as graph coloring, knapsack, and cryptarithmetic? By first-order, we mean t…
cs.CL2024
RetinaQA: A Robust Knowledge Base Question Answering Model for both Answerable and Unanswerable Questions
Prayushi Faldu, Indrajit Bhattacharya, Mausam
An essential requirement for a real-world Knowledge Base Question Answering (KBQA) system is the ability to detect the answerability of questions when generating logical forms. How…