3 papers
cs.CL2025
AMBEDKAR-A Multi-level Bias Elimination through a Decoding Approach with Knowledge Augmentation for Robust Constitutional Alignment of Language Models
Snehasis Mukhopadhyay, Aryan Kasat, Shivam Dubey +5
Large Language Models (LLMs) can inadvertently reflect societal biases present in their training data, leading to harmful or prejudiced outputs. In the Indian context, our empirica…
cs.CL2025
HumorPlanSearch: Structured Planning and HuCoT for Contextual AI Humor
Shivam Dubey
Automated humor generation with Large Language Models (LLMs) often yields jokes that feel generic, repetitive, or tone-deaf because humor is deeply situated and hinges on the liste…
cs.AI2025
Activation Steering for Bias Mitigation: An Interpretable Approach to Safer LLMs
Shivam Dubey
As large language models (LLMs) become more integrated into societal systems, the risk of them perpetuating and amplifying harmful biases becomes a critical safety concern. Traditi…