3 papers
cs.AI2026
Teaching an Agent to Sketch One Part at a Time
Xiaodan Du, Ruize Xu, David Yunis +2
We develop a method for producing vector sketches one part at a time. To do this, we train a multi-modal language model-based agent using a novel multi-turn process-reward reinforc…
cs.CL2025
SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction
Shester Gueuwou, Xiaodan Du, Greg Shakhnarovich +2
Sign language processing has traditionally relied on task-specific models, limiting the potential for transfer learning across tasks. Pre-training methods for sign language have ty…
cs.CL2025
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
Shester Gueuwou, Xiaodan Du, Greg Shakhnarovich +1
A persistent challenge in sign language video processing, including the task of sign to written language translation, is how we learn representations of sign language in an effecti…