Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
AIDABench: AI Data Analytics Benchmark
Yibo Yang, Fei Lei, Yixuan Sun +24
As AI-driven document understanding and processing tools become increasingly prevalent in real-world applications, the need for rigorous evaluation standards has grown increasingly…
cs.AI2025
Thinking by Doing: Building Efficient World Model Reasoning in LLMs via Multi-turn Interaction
Bao Shu, Yan Cai, Jianjian Sun +11
Developing robust world model reasoning is crucial for large language model (LLM) agents to plan and interact in complex environments. While multi-turn interaction offers a superio…
cs.AI2025
Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs
Chang Li, Yaren Zhang, Haoran Lv +3
Large Language Models (LLMs) have shown remarkable reasoning ability through explicit Chain-of-Thought (CoT) prompting, but generating these step-by-step textual explanations is co…