3 papers
cs.CL2026
Stability-Aware Feature Design for Robust Watermark Detection in Machine-Generated Text
Sina Mansouri, Mohit Marvania, Abolfazl Safikhani
The widespread adoption of large language models (LLMs) has intensified the demand for principled methods to distinguish human from machine-generated text. Watermarking provides a…
cs.AI2026
VeriSim: A Configurable Framework for Evaluating Medical AI Under Realistic Patient Noise
Sina Mansouri, Mohit Marvania, Vibhavari Ashok Shihorkar +5
Medical large language models (LLMs) achieve impressive performance on standardized benchmarks, yet these evaluations fail to capture the complexity of real clinical encounters whe…
cs.CL2026
Batch Prompting Suppresses Overthinking Reasoning Under Constraint: How Batch Prompting Suppresses Overthinking in Reasoning Models
Saurabh Srivastava, Janit Bidhan, Hao Yan +7
Large Reasoning Models (LRMs) achieve strong performance through explicit chain-of-thought reasoning but suffer from \textit{overthinking}: generating excessive reasoning tokens ev…