Showing 2025Show all
3 papers · 1 filter
cs.LG2025
Quantifying Robustness: A Benchmarking Framework for Deep Learning Forecasting in Cyber-Physical Systems
Alexander Windmann, Henrik Steude, Daniel Boschmann +1
Cyber-Physical Systems (CPS) in domains such as manufacturing and energy distribution generate complex time series data crucial for Prognostics and Health Management (PHM). While D…
cs.LG2025
MAWIFlow Benchmark: Realistic Flow-Based Evaluation for Network Intrusion Detection
Joshua Schraven, Alexander Windmann, Oliver Niggemann
Benchmark datasets for network intrusion detection commonly rely on synthetically generated traffic, which fails to reflect the statistical variability and temporal drift encounter…
cs.AI2025
Evaluating Large Language Models for Real-World Engineering Tasks
Rene Heesch, Sebastian Eilermann, Alexander Windmann +3
Large Language Models (LLMs) are transformative not only for daily activities but also for engineering tasks. However, current evaluations of LLMs in engineering exhibit two critic…