Showing 2026Show all
3 papers · 1 filter
cs.CL2026
DynaWeb: Model-Based Reinforcement Learning of Web Agents
Hang Ding, Peidong Liu, Junqiao Wang +7
The development of autonomous web agents, powered by Large Language Models (LLMs) and reinforcement learning (RL), represents a significant step towards general-purpose AI assistan…
cs.CV2026
iSight: Towards expert-AI co-assessment for improved immunohistochemistry staining interpretation
Jacob S. Leiby, Jialu Yao, Pan Lu +17
Immunohistochemistry (IHC) provides information on protein expression in tissue sections and is commonly used to support pathology diagnosis and disease triage. While AI models for…
cs.AI2026
STEER: Inference-Time Risk Control via Constrained Quality-Diversity Search
Eric Yang, Jong Ha Lee, Jonathan Amar +2
Large Language Models (LLMs) trained for average correctness often exhibit mode collapse, producing narrow decision behaviors on tasks where multiple responses may be reasonable. T…