3 papers
cs.AI2026
Where Do CoT Training Gains Land in LLM based Agents?
Jingyu Liu, Zhiwen Wang, Yuxin Jing +2
Chain-of-thought (CoT) reasoning is widely used in language-model agents, but prior work has shown that verbalized CoT is not always faithful and may instead reflect post-hoc reaso…
cs.IR2026
ChartWalker: Benchmarking the Cross-Chart RAG Task with Hierarchical Knowledge Graphs
Ning Tang, Chenghan Xie, Hanyang Yuan +6
Cross-Chart Retrieval-Augmented Generation (RAG) is critical for complex multi-modal analytical tasks in scientific, business, and political domains. However, existing benchmarks e…
cs.CL2025
Seed1.5-Thinking: Advancing Superb Reasoning Models with Reinforcement Learning
ByteDance Seed, :, Jiaze Chen +267
We introduce Seed1.5-Thinking, capable of reasoning through thinking before responding, resulting in improved performance on a wide range of benchmarks. Seed1.5-Thinking achieves 8…