Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Process Supervision of Confidence Margin for Calibrated LLM Reasoning
Liaoyaqi Wang, Chunsheng Zuo, William Jurayj +2
Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability. Yet, outcome-based reward of…
cs.LG2025
TAU: Modeling Temporal Consistency Through Temporal Attentive U-Net for PPG Peak Detection
Chunsheng Zuo, Yu Zhao, Juntao Ye
Photoplethysmography (PPG) sensors have been widely used in consumer wearable devices to monitor heart rates (HR) and heart rate variability (HRV). Despite the prevalence, PPG sign…
cs.LG2024
Breaking Symmetry When Training Transformers
Chunsheng Zuo, Michael Guerzhoy
As we show in this paper, the prediction for output token of Transformer architectures without one of the mechanisms of positional encodings and causal attention is invariant…