◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

H. Tamba

3 papers hereh-index 3271 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author3

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.SE1

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.LG2026

Position Bias is Hidden Behind Ceiling Effects: A Permutation Diagnostic for LLM Benchmarks

Hiroki Tamba

Position bias in multiple-choice LLM evaluation is widely cited as a confound in capability comparisons, but published measurements rely on single answer-order shuffles whose resul…

cs.SE2026

Compaction as Epistemic Failure: How Agentic LLM Tools Fabricate Confirmed Results from Killed Processes

Hiroki Tamba

Agentic LLM coding tools compress long session histories into compaction summaries that subsequent sessions inherit as ground truth. This paper documents a failure mode in Claude C…

cs.LG2026

Necessary but Not Sufficient: Temperature Control and Reproducibility in LLM-as-Judge Safety Evaluations

Hiroki Tamba

LLM-as-judge ("grader") components are now standard in evaluation harnesses, including safety evaluations where a pass/fail verdict may gate downstream deployment decisions. A wide…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.