2 citations · 3 across the 2 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024
Subobject-level Image Tokenization
Delong Chen, Samuel Cahyawijaya, Jianfeng Liu +2
Patch-based image tokenization ignores the morphology of the visual world, limiting effective and efficient learning of image understanding. Inspired by subword tokenization, we in…
cs.CV2023★ 2 cited
A Unified Framework for Multimodal, Multi-Part Human Motion Synthesis
Zixiang Zhou, Yu Wan, Baoyuan Wang
The field has made significant progress in synthesizing realistic human motion driven by various modalities. Yet, the need for different methods to animate various body parts accor…
cs.CV2023★ 1 cited
AvatarGPT: All-in-One Framework for Motion Understanding, Planning, Generation and Beyond
Zixiang Zhou, Yu Wan, Baoyuan Wang
Large Language Models(LLMs) have shown remarkable emergent abilities in unifying almost all (if not every) NLP tasks. In the human motion-related realm, however, researchers still…