2 citations · 4 across the 4 of their papers we have counts for
4 papers
HieraTok: Multi-Scale Visual Tokenizer Improves Image Reconstruction and Generation
Cong Chen, Ziyuan Huang, Cheng Zou +6
In this work, we present HieraTok, a novel multi-scale Vision Transformer (ViT)-based tokenizer that overcomes the inherent limitation of modeling single-scale representations. Thi…
Generative Active Learning for Long-tailed Instance Segmentation
Muzhi Zhu, Chengxiang Fan, Hao Chen +4
Recently, large-scale language-image generative models have gained widespread attention and many works have utilized generated data from these models to further enhance the perform…
De novo protein design using geometric vector field networks
Weian Mao, Muzhi Zhu, Zheng Sun +4
Innovations like protein diffusion have enabled significant progress in de novo protein design, which is a vital topic in life science. These methods typically depend on protein st…
SegPrompt: Boosting Open-world Segmentation via Category-level Prompt Learning
Muzhi Zhu, Hengtao Li, Hao Chen +5
Current closed-set instance segmentation models rely on pre-defined class labels for each mask during training and evaluation, largely limiting their ability to detect novel object…