3 papers
cs.CL2026
CORE: Comprehensive Ontological Relation Evaluation for Large Language Models
Satyam Dwivedi, Sanjukta Ghosh, Shivam Dwivedi +7
Large Language Models (LLMs) perform well on many reasoning benchmarks, yet existing evaluations rarely assess their ability to distinguish between meaningful semantic relations an…
cs.CV2025
Visual Language Model as a Judge for Object Detection in Industrial Diagrams
Sanjukta Ghosh
Industrial diagrams such as piping and instrumentation diagrams (P&IDs) are essential for the design, operation, and maintenance of industrial plants. Converting these diagrams int…
cs.CL2024
Machine Generated Product Advertisements: Benchmarking LLMs Against Human Performance
Sanjukta Ghosh
This study compares the performance of AI-generated and human-written product descriptions using a multifaceted evaluation model. We analyze descriptions for 100 products generated…