Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Open Vocabulary Panoptic Segmentation With Retrieval Augmentation
Nafis Sadeq, Qingfeng Liu, Mostafa El-Khamy
Given an input image and set of class names, panoptic segmentation aims to label each pixel in an image with class labels and instance labels. In comparison, Open Vocabulary Panopt…
cs.CV2025
Hardware-Friendly Static Quantization Method for Video Diffusion Transformers
Sanghyun Yi, Qingfeng Liu, Mostafa El-Khamy
Diffusion Transformers for video generation have gained significant research interest since the impressive performance of SORA. Efficient deployment of such generative-AI models on…
cs.CV2024
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation
Qingfeng Liu, Mostafa El-Khamy, Kee-Bong Song
The third Pixel-level Video Understanding in the Wild (PVUW CVPR 2024) challenge aims to advance the state of art in video understanding through benchmarking Video Panoptic Segment…