5 papers · 1 filter
Geometry-aware Incremental Neural Operator for Long-Horizon PDE prediction
Jiaquan Zhang, Shuxu Chen, Haifan Meng +6
Neural operators have shown strong potential for learning solution operators of partial differential equations (PDEs). However, long-horizon autoregressive prediction remains chall…
CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning
Shuxu Chen, Yitian Zhou, Jiaquan Zhang +6
Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However, CoT reasoning can be unstabl…
Lightweight LLM Agent Memory with Small Language Models
Jiaquan Zhang, Chaoning Zhang, Shuxu Chen +9
Although LLM agents can leverage tools for complex tasks, they still need memory to maintain cross-turn consistency and accumulate reusable information in long-horizon interactions…
Learning Global Hypothesis Space for Enhancing Synergistic Reasoning Chain
Jiaquan Zhang, Chaoning Zhang, Shuxu Chen +9
Chain-of-Thought (CoT) has been shown to significantly improve the reasoning accuracy of large language models (LLMs) on complex tasks. However, due to the autoregressive, step-by-…
Understanding Chain-of-Thought in Large Language Models via Topological Data Analysis
Chenghao Li, Chaoning Zhang, Yi Lu +10
With the development of large language models (LLMs), particularly with the introduction of the long reasoning chain technique, the reasoning ability of LLMs in complex problem-sol…