Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Uncovering Hidden Correctness in LLM Causal Reasoning via Symbolic Verification
Paul He, Yinya Huang, Mrinmaya Sachan +1
Large language models (LLMs) are increasingly being applied to tasks that involve causal reasoning. However, current benchmarks often rely on string matching or surface-level metri…
cs.AI2026
Foundations of Global Consistency Checking with Noisy LLM Oracles
Paul He, Elke Kirschbaum, Shiva Kasiviswanathan
Ensuring that collections of natural-language facts are globally consistent is essential for tasks such as fact-checking, summarization, and knowledge base construction. While Larg…