From the 1 of 55 linked papers with an AI index.
55 papers
Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning
Kunbin Xu, Xingzuo Li, Xuefeng Bai +1
Test-time reinforcement learning (TTRL) improves the reasoning capabilities of large language models without labeled data by updating the policy with pseudo-labels constructed thro…
DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation
Hongbin Zhang, Junhao Liu, Xuefeng Bai +3
The paper introduces DualAnchor, a training framework for gloss-free sign language translation that preserves large language model priors with token-level prior anchoring and enhan…
SearchSkill: Teaching LLMs to Use Search Tools with Evolving Skill Banks
Jinchao Hu, Meizhi Zhong, Kehai Chen +1
Teaching language models to use search tools is not only a question of whether they search, but also of whether they issue good queries. This is especially important in open-domain…
Agentic Tool Use in Large Language Models
Jinchao Hu, Meizhi Zhong, Kehai Chen +2
Large language models are increasingly being deployed as autonomous agents yet their real world effectiveness depends on reliable tools for information retrieval, computation and e…
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey
Bingzheng Qu, Kehai Chen, Xuefeng Bai +1
Recent progress in multimodal large language models (MLLMs) is reshaping video translation from a cascaded pipeline of automatic speech recognition, machine translation, text-to-sp…
Beyond Rigid: Benchmarking Non-Rigid Video Editing
Bingzheng Qu, Xuefeng Bai, Kehai Chen +1
As video generation models are increasingly expected to manipulate physical dynamics, there is a growing need to move evaluation beyond appearance fidelity and semantic alignment.…