model compression 1multi-teacher aggregation 1policy optimization 1reasoning models 1reinforcement learning 1teacher-student distillation 1
From the 1 of 12 linked papers with an AI index.
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Latent Reasoning with Normalizing Flows
Guancheng Tu, Xiangjun Fu, Suhao Yu +5
Large language models often improve reasoning by generating explicit chain-of-thought (CoT), demonstrating the importance of intermediate computation. However, textual CoT forces t…
cs.CL2025
Round Attention: A Novel Round-Level Attention Mechanism to Accelerate LLM Inference
Yaohua Tang, Zhicheng Hu, Kun Cheng +4
The increasing context window size in large language models (LLMs) has improved their ability to handle complex, long-text tasks. However, as the conversation rounds continue, it i…
cs.CL2024
A Full-duplex Speech Dialogue Scheme Based On Large Language Models
Peng Wang, Songshuo Lu, Yaohua Tang +3
We present a generative dialogue system capable of operating in a full-duplex manner, allowing for seamless interaction. It is based on a large language model (LLM) carefully align…