5 papers
DLR: Zero-Inference-Cost Latent Residuals for Low-Rank Pre-Training
Dong Wang, Wenwu Tang, Yun Cheng +1
Large language models have driven recent progress in language and multimodal AI, yet pre-training them at scale is prohibitively expensive. Low-rank pre-training, which factorizes…
Cut Less, Fold More: Model Compression through the Lens of Projection Geometry
Olga Saukh, Dong Wang, Haris Å ikiÄ +2
Compressing neural networks without retraining is vital for deployment at scale. We study calibration-free compression through the lens of projection geometry: structured pruning i…
Physics-Guided Inductive Spatiotemporal Kriging for PM2.5 with Satellite Gradient Constraints
Shuo Wang, Mengfan Teng, Yun Cheng +8
High-resolution mapping of fine particulate matter (PM2.5) is a cornerstone of sustainable urbanism but remains critically hindered by the spatial sparsity of ground monitoring net…
STransformer: Scalable Structured Transformers for Global Station Weather Forecasting
Hongyi Chen, Xiucheng Li, Xinyang Chen +4
Global Station Weather Forecasting (GSWF) is a key meteorological research area, critical to energy, aviation, and agriculture. Existing time series forecasting methods often ignor…
PCDCNet: A Surrogate Model for Air Quality Forecasting with Physical-Chemical Dynamics and Constraints
Shuo Wang, Yun Cheng, Qingye Meng +6
Air quality forecasting (AQF) is critical for public health and environmental management, yet remains challenging due to the complex interplay of emissions, meteorology, and chemic…