2 papers
cs.AI2026
MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use
Mengru Wang, Haozhe Luo, Zhenqian Xu +6
Memory has become a key component of large language models, enabling them to retain information and learn from long-term interactions. However, existing memory benchmarks mainly ev…
cs.AI2026
SciAtlas: A Computable Atlas of Science for Knowledge-Grounded AI Research
Shuofei Qiao, Yunxiang Wei, Busheng Zhang +16
Artificial intelligence is rapidly entering the core workflows of scientific research. Yet reliable scientific reasoning requires access to accumulated scientific knowledge with su…