18 citations
- B. Tanatar2 profiles5 · h 5
- B. Tekin3 · h 30
- Himanshu Gupta3 profiles3 · h 10
- Micah Carroll2 profiles3 · h 15
- Michael Kirchhof2 profiles3 · h 10
- Orion Weller2 profiles3 · h 22
- T. Yildiz3 · h 2
- Zheng-Xin Yong3 profiles3 · h 18
- Abdallah Galal2 profiles2 · h 1
- Alon Amit2 profiles2 · h 8
- Ankit Singh2 profiles2 · h 3
- Antonella Pinto2 profiles2 · h 2
- Centre National de la Recherche ScientifiqueFR3 papers
- Humboldt-Universität zu BerlinDE3 papers
- Ankara UniversityTR2 papers
- ETH ZurichCH2 papers
- Johns Hopkins UniversityUS2 papers
- Korea Advanced Institute of Science and TechnologyKR2 papers
- Middle East Technical UniversityTR2 papers
- National University of SingaporeSG2 papers
- The University of TokyoJP2 papers
- TU WienAT2 papers
- University College LondonGB2 papers
- University of California, BerkeleyUS2 papers
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026★ 18 cited
Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1144
Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…
cs.LG2026
Queueing-Aware Optimization of Reasoning Tokens for Accuracy-Latency Trade-offs in LLM Servers
Emre Ozbas, Melih Bastopcu
We consider a single large language model (LLM) server that serves a heterogeneous stream of queries belonging to distinct task types. Queries arrive according to a Poisson pro…