3 papers
cs.LG2026
Behavior-Consistent Deep Reinforcement Learning
Marcel Hussing, Liv G. d'Aliberti, Claas Voelcker +2
Reinforcement learning (RL) often exhibits high variance across training runs, leading to unreliable performance and posing a major challenge to deployment in real-world domains. I…
cs.AI2026
The Illusion of Insight in Reasoning Models
Liv G. d'Aliberti, Manoel Horta Ribeiro
Do reasoning models have "Aha!" moments? Prior work suggests that models like DeepSeek-R1-Zero undergo sudden mid-trace realizations that lead to accurate outputs, implying an intr…
cs.CR2025
BlindFL: Segmented Federated Learning with Fully Homomorphic Encryption
Evan Gronberg, Liv d'Aliberti, Magnus Saebo +1
Federated learning (FL) is a popular privacy-preserving edge-to-cloud technique used for training and deploying artificial intelligence (AI) models on edge devices. FL aims to secu…