6 papers · 1 filter
Judging LLM-as-a-Judge: Concerning Rubric Artifacts in LLM-based Automated Text Generation Evaluation
Anshul Bagaria, Sowmya S Sundaram, Gokul S Krishnan +1
LLM-as-a-Judge pipelines are increasingly used to evaluate AI-generated text, based on the assumption that judgments arise from reasoning over candidate responses with respect to a…
IndiCASA: A Dataset and Bias Evaluation Framework in LLMs Using Contrastive Embedding Similarity in the Indian Context
Santhosh G S, Akshay Govind S, Gokul S Krishnan +2
Large Language Models (LLMs) have gained significant traction across critical domains owing to their impressive contextual understanding and generative capabilities. However, their…
Where Should I Study? Biased Language Models Decide! Evaluating Fairness in LMs for Academic Recommendations
Krithi Shailya, Akhilesh Kumar Mishra, Gokul S Krishnan +1
Large Language Models (LLMs) are increasingly used as daily recommendation systems for tasks like education planning, yet their recommendations risk perpetuating societal biases. T…
LExT: Towards Evaluating Trustworthiness of Natural Language Explanations
Krithi Shailya, Shreya Rajpal, Gokul S Krishnan +1
As Large Language Models (LLMs) become increasingly integrated into high-stakes domains, there have been several approaches proposed toward generating natural language explanations…
InSaAF: Incorporating Safety through Accuracy and Fairness | Are LLMs ready for the Indian Legal Domain?
Yogesh Tripathi, Raghav Donakanti, Sahil Girhepuje +7
Recent advancements in language technology and Artificial Intelligence have resulted in numerous Language Models being proposed to perform various tasks in the legal domain ranging…
LineConGraphs: Line Conversation Graphs for Effective Emotion Recognition using Graph Neural Networks
Gokul S Krishnan, Sarala Padi, Craig S. Greenberg +3
Emotion Recognition in Conversations (ERC) is a critical aspect of affective computing, and it has many practical applications in healthcare, education, chatbots, and social media…