◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Sam Miller

2 papers hereh-index 3161 citations17 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author1

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.AI1
  • cs.CL1

identity via Semantic Scholar / OpenAlex

most citedWorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

1 citations · 1 across the 2 of their papers we have counts for

collaborators

2 papers

cs.AI2026

WorkBench Revisited: Workplace Agents Two Years On

Olly Styles, Sam Miller

The best agent on WorkBench in March 2024, GPT-4, completed just 43% of tasks. We revisit the benchmark in June 2026 and find that the best agent to date, Claude Fable 5, now compl…

cs.CL2024★ 1 cited

WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting

Olly Styles, Sam Miller, Patricio Cerda-Mardini +3

We introduce WorkBench: a benchmark dataset for evaluating agents' ability to execute tasks in a workplace setting. WorkBench contains a sandbox environment with five databases, 26…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.