2 papers
cs.AR2026
DPU or GPU for Accelerating Neural Networks Inference -- Why not both? Split CNN Inference
Ali Emre Oztas, Mahir Demir, James Garside +1
Video and image streaming on edge devices requires low latency. To address this, Neural Networks (NNs) are widely used, and prior work mainly focuses on accelerating them with sing…
cs.AI2024
Agentic-HLS: An agentic reasoning based high-level synthesis system using large language models (AI for EDA workshop 2024)
Ali Emre Oztas, Mahdi Jelodari
Our aim for the ML Contest for Chip Design with HLS 2024 was to predict the validity, running latency in the form of cycle counts, utilization rate of BRAM (util-BRAM), utilization…