◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Patrick Wilhelm

7 papers hereh-index 342 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author6
  • middle author1

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.CL2
  • cs.AI1
  • cs.DC1

identity via Semantic Scholar / OpenAlex

works on
activation analysis 1context calibration 1foundation models 1LLM agents 1policy size 1post-training compute allocation 1reinforcement learning 1reward feedback 1reward hacking 1safety monitoring 1search rollouts 1

From the 2 of 7 linked papers with an AI index.

collaborators
Showing cs.CLShow all

2 papers · 1 filter

cs.CL2026

Beyond Test-Time Compute Strategies: Advocating Energy-per-Token in LLM Inference

Patrick Wilhelm, Thorsten Wittkopp, Odej Kao

Large Language Models (LLMs) demonstrate exceptional performance across diverse tasks but come with substantial energy and computational costs, particularly in request-heavy scenar…

cs.CL2026

Monitoring Emergent Reward Hacking During Generation via Internal Activations

Patrick Wilhelm, Thorsten Wittkopp, Odej Kao

Fine-tuned large language models can exhibit reward-hacking behavior arising from emergent misalignment, which is difficult to detect from final outputs alone. While prior work has…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.