2 papers
cs.SE2026
AmbiBench: Benchmarking Mobile GUI Agents Beyond One-Shot Instructions in the Wild
Jiazheng Sun, Mingxuan Li, Yingying Zhang +11
Benchmarks are paramount for gauging progress in the domain of Mobile GUI Agents. In practical scenarios, users frequently fail to articulate precise directives containing full tas…
cs.CL2024
Seal: Advancing Speech Language Models to be Few-Shot Learners
Shuyu Lei, Lingen Liu, Jiaolong Yang +4
Existing auto-regressive language models have demonstrated a remarkable capability to perform a new task with just a few examples in prompt, without requiring any additional traini…