2 papers
cs.LG2026
ExecuTorch -- A Unified PyTorch Solution to Run AI Models On-Device
Mergen Nachin, Digant Desai, Sicheng Stephen Jia +36
Local execution of AI on edge devices is important for low latency and offline operation. However, deploying models on diverse hardware remains fragmented, often requiring model co…
cs.DC2024
Llama Guard 3-1B-INT4: Compact and Efficient Safeguard for Human-AI Conversations
Igor Fedorov, Kate Plawiak, Lemeng Wu +17
This paper presents Llama Guard 3-1B-INT4, a compact and efficient Llama Guard model, which has been open-sourced to the community during Meta Connect 2024. We demonstrate that Lla…