3 papers
cs.CL2026
Scaling Natural-Language Graph-Based Test Time Compute for Automated Theorem Proving
Vincent Li, Tim Knappe, Yule Fu +2
Large language models have demonstrated remarkable capabilities in natural language processing tasks requiring multi-step logical reasoning capabilities, such as automated theorem…
cs.SE2026
TDFlow: Agentic Workflows for Test Driven Development
Kevin Han, Siddharth Maddikayala, Tim Knappe +3
We introduce TDFlow, a novel test-driven agentic workflow that frames repository-scale software engineering as a test-resolution task, specifically designed to solve human-written…
cs.CL2025
Semantic Self-Consistency: Enhancing Language Model Reasoning via Semantic Weighting
Tim Knappe, Ryan Li, Ayush Chauhan +3
While large language models (LLMs) have rapidly improved their performance on a broad number of tasks, they still often fall short on reasoning tasks. As LLMs become more integrate…