From the 1 of 23 linked papers with an AI index.
1 citations · 1 across the 10 of their papers we have counts for
3 papers · 1 filter
AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment
Jianfei Xiao, Xiang Yu, Chengbing Wang +8
As Large Language Models (LLMs) evolve into lifelong AI assistants, LLM personalization has become a critical frontier. However, progress is currently bottlenecked by the absence o…
TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems
Chen Xu, Yicheng Hu, Ruizi Wang +4
Multi-agent systems (MAS) have emerged as a promising paradigm for solving complex tasks. Recent work has explored self-evolving MAS that automatically optimize agent capabilities…
Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowledge Discovery
Chaoqun Yang, Xinyu Lin, Shulin Li +4
Recent advancements in Large Language Model (LLM) agents have demonstrated remarkable potential in automatic knowledge discovery. However, rigorously evaluating an AI's capacity fo…