4 papers
Evaluating the Environmental Impact of using SLMs and Prompt Engineering for Code Generation
Md Afif Al Mamun, Sayan Nath, Gias Uddin +1
The shift from cloud-hosted Large Language Models (LLMs) to locally deployed open-source Small Language Models (SLMs) has democratized AI-assisted coding; however, it has also dece…
VLDBench Evaluating Multimodal Disinformation with Regulatory Alignment
Shaina Raza, Ashmal Vayani, Aditya Jain +8
Detecting disinformation that blends manipulated text and images has become increasingly challenging, as AI tools make synthetic content easy to generate and disseminate. While mos…
Bias Testing and Mitigation in Black Box LLMs using Metamorphic Relations
Sina Salimian, Gias Uddin, Sumon Biswas +1
The widespread deployment of Large Language Models (LLMs) has intensified concerns about subtle social biases embedded in their outputs. Existing guardrails often fail when faced w…
PCS: Perceived Confidence Scoring of Black Box LLMs with Metamorphic Relations
Sina Salimian, Gias Uddin, Shaina Raza +1
Zero-shot LLMs are now also used for textual classification tasks, e.g., sentiment and bias detection in a sentence or article. However, their performance can be suboptimal in such…