code generation 2large language models 2reinforcement learning 2adversarial red-teaming 1distribution matching 1generative flow networks 1knowledge distillation 1on-policy distillation 1
From the 2 of 18 linked papers with an AI index.
1 citations · 1 across the 8 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents
Reuben Tan, Baolin Peng, Zhengyuan Yang +16
Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally optimized using sparse, outcome-bas…
cs.AI2026
Statistical Estimation of Adversarial Risk in Large Language Models under Best-of-N Sampling
Mingqian Feng, Xiaodong Liu, Weiwei Yang +3
Large Language Models (LLMs) are typically evaluated for safety under single-shot or low-budget adversarial prompting, which underestimates real-world risk. In practice, attackers…