works on

From the 1 of 55 linked papers with an AI index.

activity
20242026
collaborators

55 papers

cs.CL2026

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning

Kunbin Xu, Xingzuo Li, Xuefeng Bai +1

Test-time reinforcement learning (TTRL) improves the reasoning capabilities of large language models without labeled data by updating the policy with pseudo-labels constructed thro…

cs.CL2026

DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation

Hongbin Zhang, Junhao Liu, Xuefeng Bai +3

The paper introduces DualAnchor, a training framework for gloss-free sign language translation that preserves large language model priors with token-level prior anchoring and enhan…

cs.AI2026

SearchSkill: Teaching LLMs to Use Search Tools with Evolving Skill Banks

Jinchao Hu, Meizhi Zhong, Kehai Chen +1

Teaching language models to use search tools is not only a question of whether they search, but also of whether they issue good queries. This is especially important in open-domain…

cs.CL2026

Agentic Tool Use in Large Language Models

Jinchao Hu, Meizhi Zhong, Kehai Chen +2

Large language models are increasingly being deployed as autonomous agents yet their real world effectiveness depends on reliable tools for information retrieval, computation and e…

cs.CV2026

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey

Bingzheng Qu, Kehai Chen, Xuefeng Bai +1

Recent progress in multimodal large language models (MLLMs) is reshaping video translation from a cascaded pipeline of automatic speech recognition, machine translation, text-to-sp…

cs.CV2026

Beyond Rigid: Benchmarking Non-Rigid Video Editing

Bingzheng Qu, Xuefeng Bai, Kehai Chen +1

As video generation models are increasingly expected to manipulate physical dynamics, there is a growing need to move evaluation beyond appearance fidelity and semantic alignment.…