From the 1 of 8 linked papers with an AI index.
8 papers
Conformalized Large Language Models under Configuration Shift
Yuqicheng Zhu, Jialin Yu, Lin Li +7
Conformal prediction (CP) is a distribution-free framework for uncertainty quantification that has recently been adapted to large language models (LLMs), providing prediction sets…
Memory for Large Language Models
Sining Zhoubian, Dan Zhang, Evgeny Kharlamov +1
The paper surveys and categorizes the various memory mechanisms used in large language models, proposing a taxonomy based on representation, update dynamics, and persistence to uni…
MADA-RL: Multi-Agent Debate-Aware Reinforcement Learning for Parameter-Efficient Reasoning in Compact Models
Martino M. L. Pulici, Cuong Xuan Chu, Evgeny Kharlamov +3
Large language models achieve strong reasoning performance, but often at prohibitive training cost - a challenge that is especially acute for compact models (…
AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs
Fanjin Zhang, Zhengyang Wang, Ruixuan Huang +7
Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-horizon tasks. Current tool-u…
Cross-Source Reasoning-based Correction for Author Name Disambiguation
Fanjin Zhang, Yunhe Pang, Bo Chen +4
Author name disambiguation is a critical challenge in academic search systems, often addressed through from-scratch and real-time disambiguation approaches. However, current algori…
Integrating Meta-Features with Knowledge Graph Embeddings for Meta-Learning
Antonis Klironomos, Ioannis Dasoulas, Francesco Periti +4
The vast collection of machine learning records available on the web presents a significant opportunity for meta-learning, where past experiments are leveraged to improve performan…