From the 1 of 5 linked papers with an AI index.
5 papers
Reward-Free Evolving Agents via Pairwise Validator
Minghao Liu, Yu Wang, Jiayun Wang +1
The paper introduces a reward‑free approach for self‑evolving agents by using a frozen large language model as a pairwise validator that decides which of two agent versions is bett…
Inference Time Optimization with Confidence Dynamics
Yu Wang, Minghao Liu, Jiayun Wang +3
Inference time optimization techniques, such as repeated sampling, have significantly advanced the reasoning capabilities of Large Language Models (LLMs). However, the critical rol…
UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation
Jiayun Wang, Yu Wang, Weijie Gan +2
We introduce spatially grounded contextual image generation, a controllable image generation task that reframes the conditioning paradigm. Instead of supplying a reference image an…
Dual-Prompt CLIP with Hybrid Visual Encoders for Occluded Person Re-Identification
Zhangjian Ji, Shaotong Qiao, Kai Feng +1
Occluded person re-identification focuses on matching partially visible pedestrians across multiple camera views. However, occlusions disrupt body-region cues, thereby complicating…
HumanLM: Simulating Users with State Alignment Beats Response Imitation
Shirley Wu, Evelyn Choi, Arpandeep Khatua +7
Large Language Models (LLMs) are increasingly used to simulate how specific users respond to a given context, enabling more user-centric applications that rely on user feedback. Ho…