4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.AI2025
Benchmarking World-Model Learning with Environment-Level Queries
Archana Warrier, Dat Nguyen, Michelangelo Naim +8
World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurable from observed interactions, s…
cs.IR2024★ 4 cited
Had enough of experts? Quantitative knowledge retrieval from large language models
David Selby, Kai Spriestersbach, Yuichiro Iwashita +6
Large language models (LLMs) have been extensively studied for their abilities to generate convincing natural language sequences, however their utility for quantitative information…