3 papers
cs.LG2025
Methodological Framework for Quantifying Semantic Test Coverage in RAG Systems
Noah Broestl, Adel Nasser Abdalla, Rajprakash Bale +2
Reliably determining the performance of Retrieval-Augmented Generation (RAG) systems depends on comprehensive test questions. While a proliferation of evaluation frameworks for LLM…
cs.CY2025
Documenting Deployment with Fabric: A Repository of Real-World AI Governance
Mackenzie Jorgensen, Kendall Brogle, Katherine M. Collins +10
Artificial intelligence (AI) is increasingly integrated into society, from financial services and traffic management to creative writing. Academic literature on the deployment of A…
cs.CY2025
Evaluating Intra-firm LLM Alignment Strategies in Business Contexts
Noah Broestl, Benjamin Lange, Cristina Voinea +2
Instruction-tuned Large Language Models (LLMs) are increasingly deployed as AI Assistants in firms for support in cognitive tasks. These AI assistants carry embedded perspectives w…