7 papers
Adaptive Hierarchical Representation Alliance for Multimodal Learning
Chunlei Meng, Pengbin Feng, Jacqueline J. Pang +5
Multimodal models often align language, vision, and audio in a single final-layer latent space, implicitly assuming that task-relevant evidence emerges at the same semantic depth a…
Rethinking Modality Reliability in Multimodal Sentiment Analysis with Incomplete Observations
Chunlei Meng, Jacqueline J. Pang, Pengbin Feng +3
Multimodal Sentiment Analysis (MSA) integrates text, audio, and vision to infer human affect, yet real-world multimodal observations are often incomplete. Existing methods for inco…
Second-Order Response Laws for LLM Judges: Debiased Estimation of Prompt Instability
Pengbin Feng, Chunlei Meng, Daozheng Qu +3
LLM judges are often evaluated with a single prompt and only a few repeated calls. When their verdicts vary, it remains unclear whether the variation comes from sampling noise with…
Staleness-Learning Rate Scaling Laws for Asynchronous RLHF
Jingwei Song, Haofeng Xu, Jie Xiao +8
High-throughput RLHF systems often decouple rollout generation from policy optimization, leading to the use of stale rollouts during learner updates. In this work, we study the eff…
MOSAIC: Orchestrating Collaborative Knowledge Tracing with Hierarchical Semantic Alignment
Xinjin Li, Mengyue Wang, Yuzhen Lin +4
Knowledge Tracing (KT) is important for personalized education but traditionally suffers from two key limitations: a reliance on shallow ID-based representations that neglect seman…
CACR:Reinforcing Temporal Answer Grounding in Instructional Video via Candidate-Aware Causal Reasoning
Muge Qi, Rong Fu, Pengbin Feng +7
The task of temporal answer grounding in instructional video (TAGV), which aims to locate precise video segments that respond to natural language queries, is increasingly important…