2 papers
cs.AI2026
WildSci: Advancing Scientific Reasoning from In-the-Wild Literature
Tengxiao Liu, Deepak Nathani, Zekun Li +2
Recent progress in large language model (LLM) reasoning has focused on domains like mathematics and coding, where abundant high-quality data and objective evaluation metrics are re…
cs.CL2025
FACTTRACK: Time-Aware World State Tracking in Story Outlines
Zhiheng Lyu, Kevin Yang, Lingpeng Kong +1
While accurately detecting and correcting factual contradictions in language model outputs has become increasingly important as their capabilities improve, doing so is highly chall…