◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Jack Wildman

4 papers hereh-index 239 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2

Across the 2 of 4 papers where every author was matched, so the position is known.

fields
  • cs.AI2
  • cs.CL1
  • cs.LG1

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.AIShow all

2 papers · 1 filter

cs.AI2026

Evaluating Strategic Reasoning in Forecasting Agents

Tom Liptay, Dan Schwarz, Rafael Poyiadzi +2

Forecasting benchmarks produce accuracy leaderboards but little insight into why some forecasters are more accurate than others. We introduce Bench to the Future 2 (BTF-2), 1,417 p…

cs.AI2025

Deep Research Bench: Evaluating AI Web Research Agents

FutureSearch, :, Nikos I. Bosse +7

Amongst the most common use cases of modern AI is LLM chat with web search enabled. However, no direct evaluations of the quality of web research agents exist that control for the…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.