2 papers
cs.LG2026
Self-Improvement as Coherence Optimization: A Theoretical Account
Tianyi Qiu, Ahmed Hani Ismail, Zhonghao He +1
Can language models improve their accuracy without external supervision? Methods such as debate, bootstrap, and internal coherence maximization achieve this surprising feat, even m…
cs.AI2025
Martingale Score: An Unsupervised Metric for Bayesian Rationality in LLM Reasoning
Zhonghao He, Tianyi Qiu, Hirokazu Shirado +1
Recent advances in reasoning techniques have substantially improved the performance of large language models (LLMs), raising expectations for their ability to provide accurate, tru…