1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2023★ 1 cited
WorldSense: A Synthetic Benchmark for Grounded Reasoning in Large Language Models
Youssef Benchekroun, Megi Dervishi, Mark Ibrahim +7
We propose WorldSense, a benchmark designed to assess the extent to which LLMs are consistently able to sustain tacit world models, by testing how they draw simple inferences from…
cs.IT2021
Near-Optimal Pool Testing under Urgency Constraints
Éric Brier, Megi Dervishi, Rémi Géraud-Stewart +2
Detection of rare traits or diseases in a large population is challenging. Pool testing allows covering larger swathes of population at a reduced cost, while simplifying logistics.…