3 papers
cs.IT2026
Probability of super-regular matrices and MDS codes over finite fields
Rathinakumar Appuswamy, Marco Bazzani, Spencer Congero +3
Let be an linear code chosen uniformly at random over a finite field of size . The following asymptotic probability of being maximum distance sepa…
cs.DC2025
A Scalable NorthPole System with End-to-End Vertical Integration for Low-Latency and Energy-Efficient LLM Inference
Michael V. DeBole, Rathinakumar Appuswamy, Neil McGlohon +30
A vertically integrated, end-to-end, research prototype system combines 288 NorthPole neural inference accelerator cards, offline training algorithms, a high-performance runtime st…
cs.LG2025
SiLQ: Simple Large Language Model Quantization-Aware Training
Steven K. Esser, Jeffrey L. McKinstry, Deepika Bablani +2
Large language models can be quantized to reduce inference time latency, model size, and energy consumption, thereby delivering a better user experience at lower cost. A challenge…