activity
20242026
collaborators

6 papers

cs.LG2026

SR-OPSD: Self-Referenced On-Policy Self-Distillation

Zhuo Sun, Entong Li, Yanlong Zhao +7

On-policy self-distillation (OPSD) converts feedback into dense token-level supervision on trajectories generated by the policy to be optimized, providing a useful complement to re…

stat.ME2026

Estimation of Directed Acyclic Graphs by Frequentist Model Averaging

Huihang Liu, Wenhui Li, Xinyu Zhang

Directed acyclic graphs provide a fundamental tool for representing directed dependence structures in multivariate network data, and are widely used to model financial and economic…

cs.AI2026

Saliency-Aware Regularized Quantization Calibration for Large Language Models

Yanlong Zhao, Xiaoyuan Cheng, Huihang Liu +6

Post-training quantization (PTQ) is an effective approach for deploying large language models (LLMs) under memory and latency constraints. Most existing PTQ methods determine quant…

stat.ME2025

Sufficiency-principled Transfer Learning via Model Averaging

Xiyuan Zhang, Huihang Liu, Xinyu Zhang

When the transferable set is unknowable, transfering informative knowledge as much as possible\textemdash a principle we refer to as \emph{sufficiency}, becomes crucial for enhanci…

stat.ML2024

Deep Generative Demand Learning for Newsvendor and Pricing

Shijin Gong, Huihang Liu, Xinyu Zhang

We consider data-driven inventory and pricing decisions in the feature-based newsvendor problem, where demand is influenced by both price and contextual features and is modeled wit…

stat.ME2024

Semi-supervised learning using copula-based regression and model averaging

Ziwen Gao, Huihang Liu, Xinyu Zhang

The available data in semi-supervised learning usually consists of relatively small sized labeled data and much larger sized unlabeled data. How to effectively exploit unlabeled da…