2 papers
cs.MA2026
Safe Equilibrium Policy Optimization for Strategic Agent Policies
Karthika Arumugam, Kiran Kumar Manku, Amit Dhanda
Language models fine-tuned with reinforcement learning typically optimize for task reward, ignoring multi-agent strategic structure. Because these agents condition on natural langu…
cs.AI2025
The Amazon Nova Family of Models: Technical Report and Model Card
Amazon AGI, Aaron Langford, Aayush Shah +783
We present Amazon Nova, a new generation of state-of-the-art foundation models that deliver frontier intelligence and industry-leading price performance. Amazon Nova Pro is a highl…