From the 1 of 3 linked papers with an AI index.
3 papers
Can Argus Judge Them All? Comparing VLMs Across Domains
Harsh Joshi, Gautam Siddharth Kashyap, Rafiq Ali +5
The paper introduces ARGUS-EVAL, a framework that assesses vision-language models on both capability and reliability across domains, and uses it to compare several VLMs on retrieva…
DocSplit: A Comprehensive Benchmark Dataset and Evaluation Approach for Document Packet Recognition and Splitting
Md Mofijul Islam, Md Sirajus Salekin, Nivedha Balakrishnan +6
Document understanding in real-world applications often requires processing heterogeneous, multi-page document packets containing multiple documents stitched together. Despite rece…
Revealing the Truth with ConLLM for Detecting Multi-Modal Deepfakes
Gautam Siddharth Kashyap, Harsh Joshi, Niharika Jain +4
The rapid rise of deepfake technology poses a severe threat to social and political stability by enabling hyper-realistic synthetic media capable of manipulating public perception.…