1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2025
SparAMX: Accelerating Compressed LLMs Token Generation on AMX-powered CPUs
Ahmed F. AbouElhamayed, Jordan Dotzel, Yash Akhauri +6
Large language models have high compute, latency, and memory requirements. While specialized accelerators such as GPUs and TPUs typically run these workloads, CPUs are more widely…
cs.CR2025★ 1 cited
Enhancing Data Integrity through Provenance Tracking in Semantic Web Frameworks
Nilesh Jain
This paper explores the integration of provenance tracking systems within the context of Semantic Web technologies to enhance data integrity in diverse operational environments. SU…