From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
SignLlama: Enhancing Gloss-free Sign Language Translation by Prioritizing Visual Features for LLMs
Shiwei Gan, Xiao Liu, Yafeng Yin +5
Large Language Models (LLMs) have achieved remarkable success across a wide range of tasks. However, fine-tuning LLMs for Gloss-Free Sign Language Translation (GFSLT) remains a cha…
cs.AI2026
Sign Language Question Answering: A New Task, Benchmark, and Baseline for Sign Language Understanding
Shiwei Gan, Lichen Wang, Xiao Liu +4
The paper introduces Sign Language Question Answering (SLQA), a task where models answer natural language questions about sign language videos, and provides two benchmark datasets…
cs.CV2026
InfoMerge: Information-aware Token Compression for Efficient Video Large Language Models
Xinxin Liu, Shiwei Gan, Xiao Liu +3
Video Large Language Models (Video-LLMs) achieve strong performance in video understanding, but their excessive visual tokens bring substantial computational overhead. Existing tra…