3 papers
cs.CL2026
Scaling Natural-Language Graph-Based Test Time Compute for Automated Theorem Proving
Vincent Li, Tim Knappe, Yule Fu +2
Large language models have demonstrated remarkable capabilities in natural language processing tasks requiring multi-step logical reasoning capabilities, such as automated theorem…
cs.SE2026
TDFlow: Agentic Workflows for Test Driven Development
Kevin Han, Siddharth Maddikayala, Tim Knappe +3
We introduce TDFlow, a novel test-driven agentic workflow that frames repository-scale software engineering as a test-resolution task, specifically designed to solve human-written…
cs.CL2024
Pragmatic Metacognitive Prompting Improves LLM Performance on Sarcasm Detection
Joshua Lee, Wyatt Fong, Alexander Le +3
Sarcasm detection is a significant challenge in sentiment analysis due to the nuanced and context-dependent nature of verbiage. We introduce Pragmatic Metacognitive Prompting (PMP)…