1.8k citations
- Zhejiang UniversityCN57 papers
- Peking UniversityCN29 papers
- Alibaba Group (United States)US28 papers
- Tsinghua UniversityCN26 papers
- Shanghai Jiao Tong UniversityCN21 papers
- University of Science and Technology of ChinaCN18 papers
- Chinese Academy of SciencesCN15 papers
- Wuhan UniversityCN14 papers
- Nanyang Technological UniversitySG12 papers
- University of Chinese Academy of SciencesCN11 papers
- Hong Kong University of Science and TechnologyHK10 papers
- Huazhong University of Science and TechnologyCN10 papers
7 papers · 2 filters
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
Xiaoguang Han, Yushuang Wu, Luyue Shi +7
MVImgNet is a large-scale dataset that contains multi-view images of ~220k real-world objects in 238 classes. As a counterpart of ImageNet, it introduces 3D visual signals via mult…
AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models
Yongjian Wu, Yang Zhou, Jiya Saiyin +4
Large-scale visual-language pre-trained models (VLPMs) have demonstrated exceptional performance in downstream object detection through text prompts for natural scenes. However, th…
GaussianTalker: Speaker-specific Talking Head Synthesis via 3D Gaussian Splatting
Hongyun Yu, Zhan Qu, Qihang Yu +8
Recent works on audio-driven talking head synthesis using Neural Radiance Fields (NeRF) have achieved impressive results. However, due to inadequate pose and expression control cau…
HierCode: A Lightweight Hierarchical Codebook for Zero-shot Chinese Text Recognition
Yuyi Zhang, Yuanzhi Zhu, Dezhi Peng +5
Text recognition, especially for complex scripts like Chinese, faces unique challenges due to its intricate character structures and vast vocabulary. Traditional one-hot encoding m…
Exploring Dynamic Transformer for Efficient Object Tracking
Jiawen Zhu, Xin Chen, Haiwen Diao +6
The speed-precision trade-off is a critical problem for visual object tracking which usually requires low latency and deployment on constrained resources. Existing solutions for ef…
Enhancing Hyperspectral Images via Diffusion Model and Group-Autoencoder Super-resolution Network
Zhaoyang Wang, Dongyang Li, Mingyang Zhang +2
Existing hyperspectral image (HSI) super-resolution (SR) methods struggle to effectively capture the complex spectral-spatial relationships and low-level details, while diffusion m…