3 papers
cs.AI2026
Reward-Oracle MCTS for Formal Theorem Proving: Sample-Efficient Search and the Need for Kernel-Level Proof Auditing
Bodla Krishna Vamshi, Haizhao Yang
Formal theorem proving with large language models remains challenging due to the difficulty of navigating large proof search spaces efficiently. Existing tree search approaches eit…
cs.LG2026
Principal Prototype Analysis on Manifold for Interpretable Reinforcement Learning
Bodla Krishna Vamshi, Haizhao Yang
Recent years have witnessed the widespread adoption of reinforcement learning (RL), from solving real-time games to fine-tuning large language models using human preference data si…
cs.LG2026
Geometry-Aware Hallucination Detection in Large Language Models
Bodla Krishna Vamshi, Rohan Bhatnagar, Haizhao Yang
Large language models (LLMs) frequently generate factually incorrect or unsupported content, commonly referred to as hallucinations. Prior work has explored decoding strategies, re…