4 papers
Heterogeneity-Aware Microscaling for Efficient Low-Bit LLM Inference
Junyi Luo, Xinting Jiang, Tai-Hao Wen +9
Microscaling (MX) is now the standard for low-bit large language model (LLM) inference. Its 4-bit form MXFP4 still loses substantial accuracy, because existing MX formats fix eithe…
CryoZip: An Efficient Cryogenic Compressor for Quantum Error Correction Syndromes
Guanchen Tao, Alexander Knapen, Jacob Mack +4
Scaling fault tolerant quantum computing is increasingly constrained by the limited bandwidth and power budget across the 4 K to room temperature (RT) interface. We present CryoZip…
Mitigating Classical Resource Costs in Quantum Error Correction via Generalized qLDPC Predecoding
Alexander Knapen, Junyi Luo, Guanchen Tao +6
Large-scale fault-tolerant quantum computing (FTQC) will require quantum-classical interfaces (QCIs) that orchestrate real-time decoding over thousands to millions of logical qubit…
Pinball: A Cryogenic Predecoder for Surface Code Decoding Under Circuit-Level Noise
Alexander Knapen, Guanchen Tao, Jacob Mack +5
Scaling fault tolerant quantum computers, especially cryogenic systems based on the surface code, to millions of qubits is challenging due to poorly-scaling data processing and pow…