7 papers
From Hypothesis to Premises: LLM-based Backward Logical Reasoning with Selective Symbolic Translation
Qingchuan Li, Mingyue Cheng, Zirui Liu +3
Logical reasoning is a core challenge in natural language understanding and a fundamental capability of artificial intelligence, underpinning scientific discovery, mathematical the…
Towards Stable and Structured Time Series Generation with Perturbation-Aware Flow Matching
Jintao Zhang, Mingyue Cheng, Zirui Liu +3
Time series generation is critical for a wide range of applications, which greatly supports downstream analytical and decision-making tasks. However, the inherent temporal heteroge…
PaperArena: An Evaluation Benchmark for Tool-Augmented Agentic Reasoning on Scientific Literature
Daoyu Wang, Mingyue Cheng, Shuo Yu +4
Understanding and reasoning on the large-scale scientific literature is a crucial touchstone for large language model (LLM) based agents. However, existing works are mainly restric…
MemWeaver: A Hierarchical Memory from Textual Interactive Behaviors for Personalized Generation
Shuo Yu, Mingyue Cheng, Daoyu Wang +4
The primary form of user-internet engagement is shifting from leveraging implicit feedback signals, such as browsing and clicks, to harnessing the rich explicit feedback provided b…
Are LLMs Stable Formal Logic Translators in Logical Reasoning Across Linguistically Diversified Texts?
Qingchuan Li, Jiatong Li, Zirui Liu +4
Logical reasoning with large language models (LLMs) has received growing attention. One mainstream approach translates natural language into formal logic and then applies symbolic…
am-ELO: A Stable Framework for Arena-based LLM Evaluation
Zirui Liu, Jiatong Li, Yan Zhuang +5
Arena-based evaluation is a fundamental yet significant evaluation paradigm for modern AI models, especially large language models (LLMs). Existing framework based on ELO rating sy…