2 papers
cs.AI2026
Verifiable Memory: Learning Unified Memory Management with Local and Global Verifiers for Large Language Model Agents
Xiaolong Sun, Qichao Wang, Hangyu Li +1
Large language model (LLM) agents must retain reusable information, control a bounded active context, and recover earlier evidence during long-horizon interaction. Existing methods…
cs.CL2025
SciDA: Scientific Dynamic Assessor of LLMs
Junting Zhou, Tingjia Miao, Yiyan Liao +15
Advancement in Large Language Models (LLMs) reasoning capabilities enables them to solve scientific problems with enhanced efficacy. Thereby, a high-quality benchmark for comprehen…