8 papers
Strategic Network Abandonment
Sandro Claudio Lera, Andreas Haupt
Socio-economic networks, from cities and firms to collaborative projects, often appear resilient for long periods before experiencing rapid, cascading decline as participation erod…
The Collapse of Heterogeneity in Silicon Philosophers
Yuanming Shi, Andreas Haupt
Silicon samples are increasingly used as a low-cost substitute for human panels and have been shown to reproduce aggregate human opinion with high fidelity. We show that, in the al…
Don't Walk the Line: Boundary Guidance for Filtered Generation
Sarah Ball, Andreas Haupt
Generative models are increasingly paired with safety classifiers that filter harmful or undesirable outputs. A common strategy is to fine-tune the generator to reduce the probabil…
Latent Adversarial Regularization for Offline Preference Optimization
Enyi Jiang, Yibo Jacky Zhang, Yinglun Xu +3
Learning from human feedback typically relies on preference optimization that constrains policy updates through token-level regularization. However, preference optimization for lan…
Full-Stack Alignment: Co-Aligning AI and Institutions with Thick Models of Value
Joe Edelman, Tan Zhi-Xuan, Ryan Lowe +30
Beneficial societal outcomes cannot be guaranteed by aligning individual AI systems with the intentions of their operators or users. Even an AI system that is perfectly aligned to…
Preference Measurement Error, Concentration in Recommendation Systems, and Persuasion
Andreas Haupt
Algorithmic recommendation based on noisy preference measurement is prevalent in recommendation systems. This paper discusses the consequences of such recommendation on market conc…