evolutionary algorithms 1large language models 1pairwise validation 1reward-free learning 1self-evolving agents 1
From the 1 of 6 linked papers with an AI index.
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Inference Time Optimization with Confidence Dynamics
Yu Wang, Minghao Liu, Jiayun Wang +3
Inference time optimization techniques, such as repeated sampling, have significantly advanced the reasoning capabilities of Large Language Models (LLMs). However, the critical rol…
cs.CL2026
HumanLM: Simulating Users with State Alignment Beats Response Imitation
Shirley Wu, Evelyn Choi, Arpandeep Khatua +7
Large Language Models (LLMs) are increasingly used to simulate how specific users respond to a given context, enabling more user-centric applications that rely on user feedback. Ho…
cs.CL2024
Gemma 2: Improving Open Language Models at a Practical Size
Gemma Team, Morgane Riviere, Shreya Pathak +195
In this work, we introduce Gemma 2, a new addition to the Gemma family of lightweight, state-of-the-art open models, ranging in scale from 2 billion to 27 billion parameters. In th…