3 papers
cs.PL2026
s2n-bignum-bench: A practical benchmark for evaluating low-level code reasoning of LLMs
Balaji Rao, John Harrison, Soonho Kong +2
Neurosymbolic approaches leveraging Large Language Models (LLMs) with formal methods have recently achieved strong results on mathematics-oriented theorem-proving benchmarks. Howev…
cs.AI2025
Neural Theorem Proving: Generating and Structuring Proofs for Formal Verification
Balaji Rao, William Eiers, Carlo Lipizzi
Formally verifying properties of software code has been a highly desirable task, especially with the emergence of LLM-generated code. In the same vein, they provide an interesting…
cs.CY2025
Steve: LLM Powered ChatBot for Career Progression
Naveen Mathews Renji, Balaji Rao, Carlo Lipizzi
The advancements in systems deploying large language models (LLMs), as well as improvements in their ability to act as agents with predefined templates, provide an opportunity to c…