4 papers
HPO: Hysteretic Policy Optimization for Stable and Efficient Training under Sparse-Reward Regime
Mohamed Sana, Nicola Piovesan, Antonio De Domenico +2
We investigate a narrow but common failure mode of GRPO-style reinforcement learning in the context of sparse verifiable rewards: early updates contain more responses with negative…
A First Measurement Study on Authentication Security in Real-World Remote MCP Servers
Huijun Zhou, Xiaohan Zhang, Haozhe Zhang +3
The Model Context Protocol (MCP) is emerging as a common interface connecting large language models (LLMs) with external services. Remote deployments are becoming increasingly impo…
TeleTables: A Benchmark for Large Language Models in Telecom Table Interpretation
Anas Ezzakri, Nicola Piovesan, Mohamed Sana +3
Language Models (LLMs) are increasingly explored in the telecom industry to support engineering tasks, accelerate troubleshooting, and assist in interpreting complex technical docu…
Reasoning Language Models for Root Cause Analysis in 5G Wireless Networks
Mohamed Sana, Nicola Piovesan, Antonio De Domenico +4
Root Cause Analysis (RCA) in mobile networks remains a challenging task due to the need for interpretability, domain expertise, and causal reasoning. In this work, we propose a lig…