3 papers
cs.SE2025
Bias Testing and Mitigation in Black Box LLMs using Metamorphic Relations
Sina Salimian, Gias Uddin, Sumon Biswas +1
The widespread deployment of Large Language Models (LLMs) has intensified concerns about subtle social biases embedded in their outputs. Existing guardrails often fail when faced w…
cs.CL2025
VLDBench Evaluating Multimodal Disinformation with Regulatory Alignment
Shaina Raza, Ashmal Vayani, Aditya Jain +8
Detecting disinformation that blends manipulated text and images has become increasingly challenging, as AI tools make synthetic content easy to generate and disseminate. While mos…
cs.CL2025
PCS: Perceived Confidence Scoring of Black Box LLMs with Metamorphic Relations
Sina Salimian, Gias Uddin, Shaina Raza +1
Zero-shot LLMs are now also used for textual classification tasks, e.g., sentiment and bias detection in a sentence or article. However, their performance can be suboptimal in such…