4 papers
Bridging Event Streams and DiT: Event-Guided Video Frame Interpolation
Guixu Lin, Yuyang Yu, Xiang Ji +6
Latent diffusion models have recently advanced video frame interpolation by synthesizing intermediate frames between input images. However, handling large temporal gaps and complex…
MoECodec: Image Compression for joint human and machine perception via Mixture-of-Experts
Jiancheng Zhao, Xiang Ji, Yifan Zhan +2
Image compression for machines calls for a unified codec that serves multiple downstream vision tasks. Existing approaches either adopt task-specific end-to-end designs, raising pa…
Tree-NeRV: A Tree-Structured Neural Representation for Efficient Non-Uniform Video Encoding
Jiancheng Zhao, Yifan Zhan, Qingtian Zhu +5
Implicit Neural Representations for Videos (NeRV) have emerged as a powerful paradigm for video representation, enabling direct mappings from frame indices to video frames. However…
All-in-One Transferring Image Compression from Human Perception to Multi-Machine Perception
Jiancheng Zhao, Xiang Ji, Yinqiang Zheng
Efficiently transferring Learned Image Compression (LIC) model from human perception to machine perception is an emerging challenge in vision-centric representation learning. Exist…