works on

From the 1 of 36 linked papers with an AI index.

activity
20242026
collaborators

36 papers

cs.CL2026

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning

Kunbin Xu, Xingzuo Li, Xuefeng Bai +1

Test-time reinforcement learning (TTRL) improves the reasoning capabilities of large language models without labeled data by updating the policy with pseudo-labels constructed thro…

cs.CL2026

DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation

Hongbin Zhang, Junhao Liu, Xuefeng Bai +3

The paper introduces DualAnchor, a training framework for gloss-free sign language translation that preserves large language model priors with token-level prior anchoring and enhan…

cs.CL2026

Agentic Tool Use in Large Language Models

Jinchao Hu, Meizhi Zhong, Kehai Chen +2

Large language models are increasingly being deployed as autonomous agents yet their real world effectiveness depends on reliable tools for information retrieval, computation and e…

cs.CV2026

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey

Bingzheng Qu, Kehai Chen, Xuefeng Bai +1

Recent progress in multimodal large language models (MLLMs) is reshaping video translation from a cascaded pipeline of automatic speech recognition, machine translation, text-to-sp…

cs.CV2026

Beyond Rigid: Benchmarking Non-Rigid Video Editing

Bingzheng Qu, Xuefeng Bai, Kehai Chen +1

As video generation models are increasingly expected to manipulate physical dynamics, there is a growing need to move evaluation beyond appearance fidelity and semantic alignment.…

cs.AI2026

Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure

Zirui Li, Xuefeng Bai, Kehai Chen +4

Latent or continuous chain-of-thought methods replace explicit textual rationales with a number of internal latent steps, but these intermediate computations are difficult to evalu…