3 papers
cs.CR2026
Multimodal Reasoning with LLM for Encrypted Traffic Interpretation: A Benchmark
Longgang Zhang, Xiaowei Fu, Fuxiang Huang +1
Network traffic, as a key media format, is crucial for ensuring security and communications in modern internet infrastructure. While existing methods offer excellent performance, t…
cs.CL2025
On-the-fly Preference Alignment via Principle-Guided Decoding
Mingye Zhu, Yi Liu, Lei Zhang +2
With the rapidly expanding landscape of large language models, aligning model generations with human values and preferences is becoming increasingly important. Popular alignment me…
cs.CL2024
LIRE: listwise reward enhancement for preference alignment
Mingye Zhu, Yi Liu, Lei Zhang +2
Recently, tremendous strides have been made to align the generation of Large Language Models (LLMs) with human values to mitigate toxic or unhelpful content. Leveraging Reinforceme…