4 papers
ConDABench: Interactive Evaluation of Language Models for Data Analysis
Avik Dutta, Priyanshu Gupta, Hosein Hasanbeig +6
Real-world data analysis tasks often come with under-specified goals and unclean data. User interaction is necessary to understand and disambiguate a user's intent, and hence, esse…
Shiksha Copilot: Teacher-AI Collaboration for Curating and Customizing Lesson Plans in Low-Resource Schools
Deepak Varuvel Dennison, Bakhtawar Ahtisham, Kavyansh Chourasia +6
This study investigates Shiksha copilot, an AI-assisted lesson planning tool deployed in government schools across Karnataka, India. The system combined LLMs and human expertise th…
Model Risk Management for Generative AI In Financial Institutions
Anwesha Bhattacharyya, Ye Yu, Hanyu Yang +4
The success of OpenAI's ChatGPT in 2023 has spurred financial enterprises into exploring Generative AI applications to reduce costs or drive revenue within different lines of busin…
Downstream bias mitigation is all you need
Arkadeep Baksi, Rahul Singh, Tarun Joshi
The advent of transformer-based architectures and large language models (LLMs) have significantly advanced the performance of natural language processing (NLP) models. Since these…