From the 1 of 6 linked papers with an AI index.
6 papers
PILA: Plug-and-Play Insertion for LLM-native Advertising
Zhaowei Zhang, Yuhan Fu, Yihang Zhang +6
The paper introduces PILA, a plug‑and‑play sidecar that rewrites LLM responses to insert sponsored content, allowing ad placement without modifying the underlying language model.
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
Zhaowei Zhang, Xiaobo Wang, Minghua Yi +5
Achieving political consensus is crucial yet challenging for the effective functioning of social governance. However, although frontier AI systems represented by large language mod…
Fusion-PSRO: Nash Policy Fusion for Policy Space Response Oracles
Jiesong Lian, Yucong Huang, Chengdong Ma +4
For solving zero-sum games involving non-transitivity, a useful approach is to maintain a policy population to approximate the Nash Equilibrium (NE). Previous studies have shown th…
Roadmap on Incentive Compatibility for AI Alignment and Governance in Sociotechnical Systems
Zhaowei Zhang, Fengshuo Bai, Mingzhi Wang +3
The burgeoning integration of artificial intelligence (AI) into human society brings forth significant implications for societal governance and safety. While considerable strides h…
Amulet: ReAlignment During Test Time for Personalized Preference Adaptation of LLMs
Zhaowei Zhang, Fengshuo Bai, Qizhi Chen +5
How to align large language models (LLMs) with user preferences from a static general dataset has been frequently studied. However, user preferences are usually personalized, chang…
RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors
Fengshuo Bai, Runze Liu, Yali Du +2
Evaluating deep reinforcement learning (DRL) agents against targeted behavior attacks is critical for assessing their robustness. These attacks aim to manipulate the victim into sp…