2 papers
cs.LG2025
LLM Misalignment via Adversarial RLHF Platforms
Erfan Entezami, Ali Naseh
Reinforcement learning has shown remarkable performance in aligning language models with human preferences, leading to the rise of attention towards developing RLHF platforms. Thes…
hep-th2025
Renormalized Volume, Polyakov Anomaly and Orbifold Riemann Surfaces
Hossein Mohammadi, Ali Naseh, Behrad Taghavi
In arXiv:2310.17536, two of the authors studied the function for orb…