2 papers
cs.CL2025
Atla Selene Mini: A General Purpose Evaluation Model
Andrei Alexandru, Antonia Calvi, Henry Broomfield +9
We introduce Atla Selene Mini, a state-of-the-art small language model-as-a-judge (SLMJ). Selene Mini is a general-purpose evaluator that outperforms the best SLMJs and GPT-4o-mini…
cs.LG2024
Toward Information Theoretic Active Inverse Reinforcement Learning
Ondrej Bajgar, Sid William Gould, Rohan Narayan Langford Mitta +3
As AI systems become increasingly autonomous, aligning their decision-making to human preferences is essential. In domains like autonomous driving or robotics, it is impossible to…