3 papers
cs.LG2025
PlatformX: An End-to-End Transferable Platform for Energy-Efficient Neural Architecture Search
Xiaolong Tu, Dawei Chen, Kyungtae Han +2
Hardware-Aware Neural Architecture Search (HW-NAS) has emerged as a powerful tool for designing efficient deep neural networks (DNNs) tailored to edge devices. However, existing me…
cs.LG2025
lm-Meter: Unveiling Runtime Inference Latency for On-Device Language Models
Haoxin Wang, Xiaolong Tu, Hongyu Ke +3
Large Language Models (LLMs) are increasingly integrated into everyday applications, but their prevalent cloud-based deployment raises growing concerns around data privacy and long…
cs.LG2025
GreenAuto: An Automated Platform for Sustainable AI Model Design on Edge Devices
Xiaolong Tu, Dawei Chen, Kyungtae Han +2
We present GreenAuto, an end-to-end automated platform designed for sustainable AI model exploration, generation, deployment, and evaluation. GreenAuto employs a Pareto front-based…