2 papers
cs.LG2026
Temperature as a Meta-Policy: Adaptive Temperature in LLM Reinforcement Learning
Haoran Dang, Cuiling Lan, Hai Wan +2
Temperature is a crucial hyperparameter in large language models (LLMs), controlling the trade-off between exploration and exploitation during text generation. High temperatures en…
cs.CR2024
LESS: Efficient Log Storage System Based on Learned Model and Minimum Attribute Tree
Zhiyang Cheng, Zizhen Zhu, Haoran Dang +2
In recent years, cyber attacks have become increasingly sophisticated and persistent. Detection and investigation based on the provenance graph can effectively mitigate cyber intru…