Showing 2024 · stat.MLShow all
2 papers · 2 filters
stat.ML2024
A Statistical Framework for Ranking LLM-Based Chatbots
Siavash Ameli, Siyuan Zhuang, Ion Stoica +1
Large language models (LLMs) have transformed natural language processing, with frameworks like Chatbot Arena providing pioneering platforms for evaluating these models. By facilit…
stat.ML2024
How many classifiers do we need?
Hyunsuk Kim, Liam Hodgkinson, Ryan Theisen +1
As performance gains through scaling data and/or model size experience diminishing returns, it is becoming increasingly popular to turn to ensembling, where the predictions of mult…