2 papers
cs.AI2026
Measuring Cross-Task Behavioral Consistency in Language Model Agents
Amritesh Banerjee, Pranil Raichura
Agent evaluation relies almost entirely on outcome metrics such as success rate, which capture whether an agent succeeds but not how consistently it behaves. We argue that behavior…
cs.CV2025
Conditional Image Synthesis with Diffusion Models: A Survey
Zheyuan Zhan, Defang Chen, Jian-Ping Mei +5
Conditional image synthesis based on user-specified requirements is a key component in creating complex visual content. In recent years, diffusion-based generative modeling has bec…