3 papers
cs.LG2026
WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation
Zihao Wu, Hongyao Tang, Yi Ma +9
Massively parallel simulation changes the data regime in which off-policy reinforcement learning (RL) is trained, challenging stabilizers designed for data-limited replay. Through…
cs.NI2026
Beyond Per-Request QoS: Coordinating Industrial Workflows with B5G/6G Network Capabilities
Qize Guo, Bjoern Riemer, Tarik Taleb +3
Beyond-5G (B5G) and 6G networks are expected to enable more complex industrial services, which often operate according to multi-phase workflows with phase-specific communication re…
cs.DC2026
KPI2KVI: A Multi Agent Workflow for Calculating Key Value Indicators from Service Descriptions
Masoud Shokrnezhad, Tarik Taleb, Yan Chen +1
Key Value Indicators (KVIs) provide a decision oriented view of a service by summarizing how operational performance translates into stakeholder value, risk, and outcomes. However,…