3 papers
cs.CL2026
BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources
Raghvendra Kumar, Devankar Raj, Sriparna Saha
India's linguistic landscape, spanning 22 scheduled languages and hundreds of marginalized dialects, has driven rapid growth in NLP datasets, benchmarks, and pretrained models. How…
cs.CV2026
When Background Matters: Breaking Medical Vision Language Models by Transferable Attack
Akash Ghosh, Subhadip Baidya, Sriparna Saha +1
Vision-Language Models (VLMs) are increasingly used in clinical diagnostics, yet their robustness to adversarial attacks remains largely unexplored, posing serious risks. Existing…
cs.CL2026
When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities
Sarmistha Das, Shreyas Guha, Suvrayan Bandyopadhyay +3
Idiomatic reasoning, deeply intertwined with metaphor and culture, remains a blind spot for contemporary language models, whose progress skews toward surface-level lexical and sema…