Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research
Hui Chen, Miao Xiong, Yujie Lu +7
Recent advancements in AI agents have demonstrated their growing potential to drive and support scientific discovery. In this work, we introduce MLR-Bench, a comprehensive benchmar…
cs.LG2024
Are Anomaly Scores Telling the Whole Story? A Benchmark for Multilevel Anomaly Detection
Tri Cao, Minh-Huy Trinh, Ailin Deng +4
Anomaly detection (AD) is a machine learning task that identifies anomalies by learning patterns from normal training data. In many real-world scenarios, anomalies vary in severity…