activity
20242026
most citedClassroom Simulacra: Building Contextual Student Generative Agents in Online Education for Learning Behavioral Simulation

12 citations · 12 across the 10 of their papers we have counts for

collaborators

11 papers

cs.AI2026

Entropy-Guided Data-Efficient Training for Multimodal Reasoning Reward Models

Shidong Yang, Tongwen Huang, Hao Wen +3

Multimodal reward models are crucial for aligning multimodal large language models with human preferences. Recent works have incorporated reasoning capabilities into these models,…

cs.CL2026

From Context to EDUs: Faithful and Structured Context Compression via Elementary Discourse Unit Decomposition

Yiqing Zhou, Yu Lei, Shuzheng Si +7

Managing extensive context remains a critical bottleneck for Large Language Models (LLMs), particularly in applications like long-document question answering and autonomous agents…

cs.CV2025

Taming Hallucinations: Boosting MLLMs' Video Understanding via Counterfactual Video Generation

Zhe Huang, Hao Wen, Aiming Hao +6

Multimodal Large Language Models (MLLMs) have made remarkable progress in video understanding. However, they suffer from a critical vulnerability: an over-reliance on language prio…

cs.LG2025

SIRI: Scaling Iterative Reinforcement Learning with Interleaved Compression

Haoming Wen, Yushi Bai, Juanzi Li +1

We introduce SIRI, Scaling Iterative Reinforcement Learning with Interleaved Compression, a simple yet effective RL approach for Large Reasoning Models (LRMs) that enables more eff…

cs.CL2025

ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute

Hao Wen, Yifan Su, Feifei Zhang +4

Recent advances in Large Language Models (LLMs) have been driven by test-time compute scaling - a strategy that improves reasoning by generating longer, sequential thought processe…

cs.AI2025

An Empirical Study of LLM Reasoning Ability Under Strict Output Length Constraint

Yi Sun, Han Wang, Jiaqiang Li +8

Recent work has demonstrated the remarkable potential of Large Language Models (LLMs) in test-time scaling. By making models think before answering, they are able to achieve much h…