3 papers
cs.LG2025
Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization
Jiayi Tian, Jinming Lu, Hai Li +4
Transformer models have achieved state-of-the-art performance across a wide range of machine learning tasks. There is growing interest in training transformers on resource-constrai…
cs.CL2025
An Empirical Study on Prompt Compression for Large Language Models
Zheng Zhang, Jinyi Li, Yihuai Lan +2
Prompt engineering enables Large Language Models (LLMs) to perform a variety of tasks. However, lengthy prompts significantly increase computational complexity and economic costs.…
cs.AI2025
DVM: Towards Controllable LLM Agents in Social Deduction Games
Zheng Zhang, Yihuai Lan, Yangsen Chen +3
Large Language Models (LLMs) have advanced the capability of game agents in social deduction games (SDGs). These games rely heavily on conversation-driven interactions and require…