5 citations · 15 across the 45 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Nondeterministic Polynomial-time Problem Challenge: An Ever-Scaling Reasoning Benchmark for LLMs
Chang Yang, Ruiyu Wang, Junzhe Jiang +9
Reasoning is the fundamental capability of large language models (LLMs). Due to the rapid progress of LLMs, there are two main issues of current benchmarks: i) these benchmarks can…
cs.AI2016
Estimating Activity at Multiple Scales using Spatial Abstractions
Majd Hawasly, Florian T. Pokorny, Subramanian Ramamoorthy
Autonomous robots operating in dynamic environments must maintain beliefs over a hypothesis space that is rich enough to represent the activities of interest at different scales. T…