3 papers
cs.AI2026
Measuring Black-Box Confidence via Reasoning Trajectories: Geometry, Coverage, and Verbalization
Marc Boubnovski Martell, Josefa Lia Stoisser, Kaspar Märtens +4
Reliable confidence estimation enables safe deployment of chain-of-thought (CoT) reasoning through text-only APIs. Yet the dominant black-box baseline, self-consistency over K samp…
cs.LG2026
MechPert: Mechanistic Consensus as an Inductive Bias for Unseen Perturbation Prediction
Marc Boubnovski Martell, Josefa Lia Stoisser, Lawrence Phillips +6
Predicting transcriptional responses to unseen genetic perturbations is essential for understanding gene regulation and prioritizing large-scale perturbation experiments. Existing…
cs.AI2025
Towards Agents That Know When They Don't Know: Uncertainty as a Control Signal for Structured Reasoning
Josefa Lia Stoisser, Marc Boubnovski Martell, Lawrence Phillips +6
Large language model (LLM) agents are increasingly deployed in structured biomedical data environments, yet they often produce fluent but overconfident outputs when reasoning over…