collaborators

13 papers

cs.CV2026

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning

Jingqi Tian, Haoji Zhang, Lin Chen +7

Chain-of-thought (CoT) reasoning can improve performance on difficult video questions but often wastes decoding tokens on simple ones. We study whether a video multimodal large lan…

cs.CV2026

ZeroSplat: Generalized Referring Segmentation in 3D Gaussian Splatting

Jiayu Ding, Meilu Song, Xiaoyi Zhang +3

Recent advancements in 3D Gaussian Splatting (3DGS) have enabled language-guided scene understanding. However, existing Referring 3D Gaussian Splatting (R3DGS) methods are fundamen…

cs.CL2026

ContextGuard: Structured Self-Auditing for Context Learning in Language Models

Hongbo Jin, Chi Wang, Haoran Tang +5

Recent benchmarks reveal that despite strong reasoning capabilities, large language models (LLMs) still struggle to faithfully apply complex contextual knowledge. These failures ar…

cs.AI2026

Context-CoT: Enhancing Context Learning via High-Quality Reasoning Synthesis

Hongbo Jin, Mingnan Zhu, Jingqi Tian +6

While LLMs excel at reasoning over prompts using static pretrained knowledge, they struggle significantly with context learning-the ability to dynamically extract, internalize, and…

cs.CV2026

VISD: Enhancing Video Reasoning via Structured Self-Distillation

Hao Lin, Kunyang Lv, Xu Jiang +5

Training VideoLLMs for complex reasoning remains challenging due to sparse sequence level rewards and the lack of fine grained credit assignment over long, temporally grounded reas…

cs.LG2026

DGPO: Distribution Guided Policy Optimization for Fine Grained Credit Assignment

Hongbo Jin, Rongpeng Zhu, Zhongjing Du +4

Reinforcement learning is crucial for aligning large language models to perform complex reasoning tasks. However, current algorithms such as Group Relative Policy Optimization suff…