154 citations · 848 across the 33 of their papers we have counts for
6 papers · 1 filter
RIGID: Recurrent GAN Inversion and Editing of Real Face Videos
Yangyang Xu, Shengfeng He, Kwan-Yee K. Wong +1
GAN inversion is indispensable for applying the powerful editability of GAN to real images. However, existing methods invert video frames individually often leading to undesired in…
MedShapeNet -- A Large-Scale Dataset of 3D Medical Shapes for Computer Vision
Jianning Li, Zongwei Zhou, Jiancheng Yang +154
Prior to the deep learning era, shape was commonly used to describe the objects. Nowadays, state-of-the-art (SOTA) algorithms in medical imaging are predominantly diverging from co…
ChiPFormer: Transferable Chip Placement via Offline Decision Transformer
Yao Lai, Jinxin Liu, Zhentao Tang +3
Placement is a critical step in modern chip design, aiming to determine the positions of circuit modules on the chip canvas. Recent works have shown that reinforcement learning (RL…
VisionLLM: Large Language Model is also an Open-Ended Decoder for Vision-Centric Tasks
Wenhai Wang, Zhe Chen, Xiaokang Chen +8
Large language models (LLMs) have notably accelerated progress towards artificial general intelligence (AGI), with their impressive zero-shot capacity for user-tailored tasks, endo…
Graph-based Topology Reasoning for Driving Scenes
Tianyu Li, Li Chen, Huijie Wang +10
Understanding the road genome is essential to realize autonomous driving. This highly intelligent problem contains two aspects - the connection relationship of lanes, and the assig…
Real-time Controllable Denoising for Image and Video
Zhaoyang Zhang, Yitong Jiang, Wenqi Shao +4
Controllable image denoising aims to generate clean samples with human perceptual priors and balance sharpness and smoothness. In traditional filter-based denoising methods, this c…