2 papers
cs.AI2025
TaskEval: Synthesised Evaluation for Foundation-Model Tasks
Dilani Widanapathiranage, Scott Barnett, Stefanus Kurniawan +1
Hallucinations are a key concern when creating applications that rely on Foundation models (FMs). Understanding where and how these subtle failures occur in an application relies o…
cs.LG2025
The M-factor: A Novel Metric for Evaluating Neural Architecture Search in Resource-Constrained Environments
Srikanth Thudumu, Hy Nguyen, Hung Du +6
Neural Architecture Search (NAS) aims to automate the design of deep neural networks. However, existing NAS techniques often focus on maximising accuracy, neglecting model efficien…