13 papers
CodecArena: Codec Quality Assessment via Visual Reinforcement Learning
Jiaye Fu, Weiqi Li, Qiankun Gao +5
Video coding is advancing into the low and ultra-low bitrate regime, driven by end-to-end codecs that replace the hand-crafted pipeline with jointly optimized neural networks and g…
Overview of Cross-Component In-loop Filters in Video Coding Standards
Zhaoyu Li, Xuewei Meng, Jiaqi Zhang +4
The paper reviews cross-component in-loop filters used in modern video coding standards, describing how they exploit luma‑chroma correlations to improve chroma quality and reduce c…
Optimized Adaptive Loop Filter in Versatile Video Coding
Xuewei Meng, Jiaqi Zhang, Chuanmin Jia +3
In the Versatile Video Coding~(VVC) standard, adaptive loop filter~(ALF), including Geometry transformation-based Adaptive Loop Filter~(GALF) and Cross Component Adaptive Loop Filt…
L2D2-GS: Learning to Densify for Feedforward Dynamic Gaussian Scene Reconstruction
Zetian Song, Chenming Wu, Junnan Liu +6
High-fidelity reconstruction of dynamic urban environments is a cornerstone of autonomous driving simulation and large-scale world modeling. While 3D Gaussian Splatting (3DGS) has…
Spark3R: Asymmetric Token Reduction Makes Fast Feed-Forward 3D Reconstruction
Zecheng Tang, Jiaye Fu, Qiankun Gao +5
Feed-forward 3D reconstruction models based on Vision Transformers can directly estimate scene geometry and camera poses from a small set of input images, but scaling them to video…
SoLAR: Error-Resilient Streamable Long-Horizon Free-Viewpoint Video Reconstruction with Anchor Activation and Latent Recalibration
Haotian Zhang, Xu Mo, Yixin Yu +7
Free-Viewpoint Video (FVV) has emerged as a cornerstone of next-generation immersive media systems and attracted widespread attention. Previous methods primarily focus on short vid…