2 papers
cs.LG2025
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs
Xilong Xie, Liang Wang, Limin Xiao +4
Large language models (LLMs) have significantly advanced the natural language processing paradigm but impose substantial demands on memory and computational resources. Quantization…
cs.AR2023
FuseFPS: Accelerating Farthest Point Sampling with Fusing KD-tree Construction for Point Clouds
Meng Han, Liang Wang, Limin Xiao +5
Point cloud analytics has become a critical workload for embedded and mobile platforms across various applications. Farthest point sampling (FPS) is a fundamental and widely used k…