2 papers
cs.LG2026
IntAttention: A Fully Integer Attention Pipeline for Efficient Edge Inference
Wanli Zhong, Haibo Feng, Zirui Zhou +2
Deploying Transformer models on edge devices is limited by latency and energy budgets. While INT8 quantization effectively accelerates the primary matrix multiplications, it expose…
cs.LG2026
RelPrism: A Multi-Faceted Pre-training Framework with Self-Generated Tasks for Relational Databases
Jinyu Yang, Cheng Yang, Junze Chen +4
Relational databases (RDBs) remain the cornerstone of modern data systems and support diverse predictive tasks. Recent relational deep learning (RDL) methods enable end-to-end pred…