1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2025
Model Consistency as a Cheap yet Predictive Proxy for LLM Elo Scores
Ashwin Ramaswamy, Nestor Demeure, Ermal Rrapaj
New large language models (LLMs) are being released every day. Some perform significantly better or worse than expected given their parameter count. Therefore, there is a need for…
cs.LG2024★ 1 cited
Towards Adapting Reinforcement Learning Agents to New Tasks: Insights from Q-Values
Ashwin Ramaswamy, Ransalu Senanayake
While contemporary reinforcement learning research and applications have embraced policy gradient methods as the panacea of solving learning problems, value-based methods can still…