activity
20142025
most citedHumanity's Last Exam

23 citations · 53 across the 7 of their papers we have counts for

collaborators

7 papers

cs.LG2025★ 23 cited

Humanity's Last Exam

Long Phan, Alice Gatti, Ziwen Han +1144

Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…

cs.AI2024

GUI Agents: A Survey

Dang Nguyen, Jian Chen, Yu Wang +27

Graphical User Interface (GUI) agents, powered by Large Foundation Models, have emerged as a transformative approach to automating human-computer interaction. These agents autonomo…

cs.CL2023★ 2 cited

PDFTriage: Question Answering over Long, Structured Documents

Jon Saad-Falcon, Joe Barrow, Alexa Siu +4

Large Language Models (LLMs) have issues with document question answering (QA) in situations where the document is unable to fit in the small context length of an LLM. To overcome…

cs.CL2023

Boosting Punctuation Restoration with Data Generation and Reinforcement Learning

Viet Dac Lai, Abel Salinas, Hao Tan +6

Punctuation restoration is an important task in automatic speech recognition (ASR) which aim to restore the syntactic structure of generated ASR texts to improve readability. While…

cs.CC2014

TrackMania is NP-complete

Franck Dernoncourt

We prove that completing an untimed, unbounded track in TrackMania Nations Forever is NP-complete by using a reduction from 3-SAT and showing that a solution can be checked in poly…

cs.HC2014★ 5 cited

Replacing the computer mouse

Franck Dernoncourt

In a few months the computer mouse will be half-a-century-old. It is known to have many drawbacks, the main ones being: loss of productivity due to constant switching between keybo…