3 papers
cs.AI2025
MathBode: Measuring the Stability of LLM Reasoning using Frequency Response
Charles L. Wang
This paper presents MathBode, a dynamic diagnostic for mathematical reasoning in large language models (LLMs). Instead of one-shot accuracy, MathBode treats each parametric problem…
cs.AI2025
MI9: An Integrated Runtime Governance Framework for Agentic AI
Charles L. Wang, Trisha Singhal, Ameya Kelkar +1
Agentic AI systems capable of reasoning, planning, and executing actions present fundamentally distinct governance challenges compared to traditional AI models. Unlike conventional…
cs.CV2025
Zebra-CoT: A Dataset for Interleaved Vision Language Reasoning
Ang Li, Charles Wang, Deqing Fu +9
Humans often use visual aids, for example diagrams or sketches, when solving complex problems. Training multimodal models to do the same, known as Visual Chain of Thought (Visual C…