1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.AI2025★ 1 cited
Scaling Reinforcement Learning for Content Moderation with Large Language Models
Hamed Firooz, Rui Liu, Yuchen Lu +15
Content moderation at scale remains one of the most pressing challenges in today's digital ecosystem, where billions of user- and AI-generated artifacts must be continuously evalua…
cs.CL2025
LLM Prompt Duel Optimizer: Efficient Label-Free Prompt Optimization
Yuanchen Wu, Saurabh Verma, Justin Lee +6
Large language models (LLMs) are highly sensitive to prompts, but most automatic prompt optimization (APO) methods assume access to ground-truth references (e.g., labeled validatio…
cs.AI2024
Towards Unified Alignment Between Agents, Humans, and Environment
Zonghan Yang, An Liu, Zijun Liu +11
The rapid progress of foundation models has led to the prosperity of autonomous agents, which leverage the universal capabilities of foundation models to conduct reasoning, decisio…