1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CL2026
One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows
Zhuochun Li, Youngmin Ko, Ali Keramati +9
Recent agent benchmarks increasingly ground evaluation in executable environments, from code repair to web navigation, app APIs, and function calling. Yet completing consequential…
stat.ML2018
Information Planning for Text Data
Vadim Smolyakov
Information planning enables faster learning with fewer training examples. It is particularly applicable when training examples are costly to obtain. This work examines the advanta…
stat.ML2018★ 1 cited
Adaptive Scan Gibbs Sampler for Large Scale Inference Problems
Vadim Smolyakov, Qiang Liu, John W. Fisher
For large scale on-line inference problems the update strategy is critical for performance. We derive an adaptive scan Gibbs sampler that optimizes the update frequency by selectin…