2 papers
cs.AI2026
The Metacognitive Bottleneck: Japanese Riddles Reveal Fundamental Limits of Machine Insight and Self-Evaluation in Reasoning AI
Masaharu Mizumoto, Dat Nguyen, Zhiheng Han +6
Benchmark saturation and training-data contamination increasingly obscure whether reported gains in large language models (LLMs) reflect genuine advances in reasoning or familiarit…
physics.chem-ph2025
Benchmarking the Impact of Active Space Selection on the VQE Pipeline for Quantum Drug Discovery
Zhi Yin, Xiaoran Li, Zhupeng Han +6
Quantum computers promise scalable treatments of electronic structure, yet applying variational quantum eigensolvers (VQE) on realistic drug-like molecules remains constrained by t…