activity
20162023
most citedFocusNetv2: Imbalanced Large and Small Organ Segmentation with Adversarial Shape Constraint for Head and Neck CT Images

96 citations · 679 across the 47 of their papers we have counts for

collaborators
Showing cs.CVShow all

55 papers · 1 filter

cs.CV2023★ 2 cited

NDC-Scene: Boost Monocular 3D Semantic Scene Completion in Normalized Device Coordinates Space

Jiawei Yao, Chuming Li, Keqiang Sun +4

Monocular 3D Semantic Scene Completion (SSC) has garnered significant attention in recent years due to its potential to predict complex semantics and geometry shapes from a single…

cs.CV2023★ 6 cited

Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following

Ziyu Guo, Renrui Zhang, Xiangyang Zhu +8

We introduce Point-Bind, a 3D multi-modality model aligning point clouds with 2D image, language, audio, and video. Guided by ImageBind, we construct a joint embedding space betwee…

cs.CV2023★ 1 cited

Omnidirectional Information Gathering for Knowledge Transfer-based Audio-Visual Navigation

Jinyu Chen, Wenguan Wang, Si Liu +2

Audio-visual navigation is an audio-targeted wayfinding task where a robot agent is entailed to travel a never-before-seen 3D environment towards the sounding source. In this artic…

cs.CV2023★ 8 cited

Denoising Diffusion Semantic Segmentation with Mask Prior Modeling

Zeqiang Lai, Yuchen Duan, Jifeng Dai +5

The evolution of semantic segmentation has long been dominated by learning more discriminative image representations for classifying each pixel. Despite the prominent advancements,…

cs.CV2023★ 1 cited

FlowFormer: A Transformer Architecture and Its Masked Cost Volume Autoencoding for Optical Flow

Zhaoyang Huang, Xiaoyu Shi, Chao Zhang +6

This paper introduces a novel transformer-based network architecture, FlowFormer, along with the Masked Cost Volume AutoEncoding (MCVA) for pretraining it to tackle the problem of…

cs.CV2023★ 1 cited

ADDP: Learning General Representations for Image Recognition and Generation with Alternating Denoising Diffusion Process

Changyao Tian, Chenxin Tao, Jifeng Dai +7

Image recognition and generation have long been developed independently of each other. With the recent trend towards general-purpose representation learning, the development of gen…