4 papers
BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese
Peilin Zhou, Bruce Leon, Xiang Ying +13
As large language models (LLMs) evolve into tool-using agents, the ability to browse the web in real-time has become a critical yardstick for measuring their reasoning and retrieva…
Tuning LLMs by RAG Principles: Towards LLM-native Memory
Jiale Wei, Shuchi Wu, Ruochen Liu +3
Memory, additional information beyond the training of large language models (LLMs), is crucial to various real-world applications, such as personal assistant. The two mainstream so…
AI-native Memory 2.0: Second Me
Jiale Wei, Xiang Ying, Tao Gao +3
Human interaction with the external world fundamentally involves the exchange of personal memory, whether with other individuals, websites, applications, or, in the future, AI agen…
Unified Mind Model: Reimagining Autonomous Agents in the LLM Era
Pengbo Hu, Xiang Ying
Large language models (LLMs) have recently demonstrated remarkable capabilities across domains, tasks, and languages (e.g., ChatGPT and GPT-4), reviving the research of general aut…