Publications (17)
Taylor saves for later: disentanglement for video prediction using Taylor representation
Ting Pan, Zhuqing Jiang, Jianan Han +3
Video prediction is a challenging task with wide application prospects in meteorology and robot systems. Existing works fail to trade off short-term and long-term prediction perfor…
Autoregressive Video Generation without Vector Quantization
Haoge Deng, Ting Pan, Haiwen Diao +6
This paper presents a novel approach that enables autoregressive video generation with high efficiency. We propose to reformulate the video generation problem as a non-quantized au…
A Survey on Role-Oriented Network Embedding
Pengfei Jiao, Xuan Guo, Ting Pan +2
Recently, Network Embedding (NE) has become one of the most attractive research topics in machine learning and data mining. NE approaches have achieved promising performance in var…
Emu3.5: Native Multimodal Models are World Learners
Yufeng Cui, Honghao Chen, Haoge Deng +20
We introduce Emu3.5, a large-scale multimodal world model that natively predicts the next state across vision and language. Emu3.5 is pre-trained end-to-end with a unified next-tok…
When LLM Rationales Become User-Facing: Effects on Trust Perception, Decision-Making, and Gaze Behaviors
Xin Sun, Ting Pan, Yajing Wang +5
Large language models (LLMs) increasingly show step-by-step reasoning rationales alongside their answers, turning reasoning from an internal model capability into a user-facing int…
Eyes Can't Always Tell: Fusing Eye Tracking and User Priors for User Modeling under AI Advice Conditions
Xin Sun, Shu Wei, Ting Pan +5
Modeling users' cognitive states (e.g., cognitive load and decision confidence) is essential for building adaptive AI in high-stakes decision-making. While eye tracking provides no…