evolutionary algorithms 1large language models 1pairwise validation 1reward-free learning 1self-evolving agents 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
Reward-Free Evolving Agents via Pairwise Validator
Minghao Liu, Yu Wang, Jiayun Wang +1
The paper introduces a reward‑free approach for self‑evolving agents by using a frozen large language model as a pairwise validator that decides which of two agent versions is bett…
cs.CL2026
Inference Time Optimization with Confidence Dynamics
Yu Wang, Minghao Liu, Jiayun Wang +3
Inference time optimization techniques, such as repeated sampling, have significantly advanced the reasoning capabilities of Large Language Models (LLMs). However, the critical rol…