collaborators

6 papers

cs.CV2025

Optimizing 3D Gaussian Splattering for Mobile GPUs

Md Musfiqur Rahman Sanim, Zhihao Shu, Bahram Afsharmanesh +5

Image-based 3D scene reconstruction, which transforms multi-view images into a structured 3D representation of the surrounding environment, is a common task across many modern appl…

cs.LG2025

Squat: Quant Small Language Models on the Edge

Xuan Shen, Peiyan Dong, Zhenglun Kong +9

A growing trend has emerged in designing high-quality Small Language Models (SLMs) with a few million parameters. This trend is driven by the increasing concerns over cloud costs,…

cs.LG2025

LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers

Xuan Shen, Zhao Song, Yufa Zhou +12

Diffusion Transformers have emerged as the preeminent models for a wide array of generative tasks, demonstrating superior performance and efficacy across various applications. The…

cs.CV2025

QuartDepth: Post-Training Quantization for Real-Time Depth Estimation on the Edge

Xuan Shen, Weize Ma, Jing Liu +9

Monocular Depth Estimation (MDE) has emerged as a pivotal task in computer vision, supporting numerous real-world applications. However, deploying accurate depth estimation models…

cs.CV2025

Open-Source Acceleration of Stable-Diffusion.cpp Deployable on All Devices

Jingxu Ng, Cheng Lv, Pu Zhao +5

Stable diffusion plays a crucial role in generating high-quality images. However, image generation is time-consuming and memory-intensive. To address this, stable-diffusion.cpp (Sd…

cs.CV2024

Fast and Memory-Efficient Video Diffusion Using Streamlined Inference

Zheng Zhan, Yushu Wu, Yifan Gong +7

The rapid progress in artificial intelligence-generated content (AIGC), especially with diffusion models, has significantly advanced development of high-quality video generation. H…