papers

Publications (55)

cs.MM2021

ANT: Learning Accurate Network Throughput for Better Adaptive Video Streaming

Jiaoyang Yin, Yiling Xu, Hao Chen +3

Adaptive Bit Rate (ABR) decision plays a crucial role for ensuring satisfactory Quality of Experience (QoE) in video streaming applications, in which past network statistics are ma…

cs.MM2024

Deep joint source-channel coding for wireless point cloud transmission

Cixiao Zhang, Mufan Liu, Wenjie Huang +3

The growing demand for high-quality point cloud transmission over wireless networks presents significant challenges, primarily due to the large data sizes and the need for efficien…

cs.CV2025

Benchmarking and Learning Multi-Dimensional Quality Evaluator for Text-to-3D Generation

Yujie Zhang, Bingyang Cui, Qi Yang +2

Text-to-3D generation has achieved remarkable progress in recent years, yet evaluating these methods remains challenging for two reasons: i) Existing benchmarks lack fine-grained e…

cs.CV2023

SJTU-TMQA: A quality assessment database for static mesh with texture map

Bingyang Cui, Qi Yang, Kaifa Yang +3

In recent years, static meshes with texture maps have become one of the most prevalent digital representations of 3D shapes in various applications, such as animation, gaming, medi…

eess.IV2025

3DGS-VBench: A Comprehensive Video Quality Evaluation Benchmark for 3DGS Compression

Yuke Xing, William Gordon, Qi Yang +3

3D Gaussian Splatting (3DGS) enables real-time novel view synthesis with high visual fidelity, but its substantial storage requirements hinder practical deployment, prompting state…

cs.CV2023

Learning Dynamic Point Cloud Compression via Hierarchical Inter-frame Block Matching

Shuting Xia, Tingyu Fan, Yiling Xu +2

3D dynamic point cloud (DPC) compression relies on mining its temporal context, which faces significant challenges due to DPC's sparsity and non-uniform structure. Existing methods…

eess.IV2024

EVAN: Evolutional Video Streaming Adaptation via Neural Representation

Mufan Liu, Le Yang, Yiling Xu +2

Adaptive bitrate (ABR) using conventional codecs cannot further modify the bitrate once a decision has been made, exhibiting limited adaptation capability. This may result in eithe…

cs.MM2018

QoE-Oriented Resource Allocation for 360-degree Video Transmission over Heterogeneous Networks

Wei Huang, Lianghui Ding, Hung-Yu Wei +3

Immersive media streaming, especially virtual reality (VR)/360-degree video streaming which is very bandwidth demanding, has become more and more popular due to the rapid growth of…

cs.CR2018

IoT Security: An End-to-End View and Case Study

Zhen Ling, Kaizheng Liu, Yiling Xu +5

In this paper, we present an end-to-end view of IoT security and privacy and a case study. Our contribution is three-fold. First, we present our end-to-end view of an IoT system an…

cs.CV2026

Rasterizing Wireless Radiance Field via Deformable 2D Gaussian Splatting

Mufan Liu, Cixiao Zhang, Qi Yang +6

Modeling the wireless radiance field (WRF) is fundamental to modern communication systems, enabling key tasks such as localization, sensing, and channel estimation. Traditional app…

cs.GR2025

Textured mesh Quality Assessment using Geometry and Color Field Similarity

Kaifa Yang, Qi Yang, Zhu Li +1

Textured mesh quality assessment (TMQA) is critical for various 3D mesh applications. However, existing TMQA methods often struggle to provide accurate and robust evaluations. Moti…

eess.IV2025

Once-Training-All-Fine: No-Reference Point Cloud Quality Assessment via Domain-relevance Degradation Description

Yipeng Liu, Qi Yang, Yujie Zhang +4

The visual quality of point clouds plays a crucial role in the development and broadcasting of immersive media. Therefore, investigating point cloud quality assessment (PCQA) is in…

cs.CV2022

3DAC: Learning Attribute Compression for Point Clouds

Guangchi Fang, Qingyong Hu, Hanyun Wang +2

We study the problem of attribute compression for large-scale unstructured 3D point clouds. Through an in-depth exploration of the relationships between different encoding steps an…

cs.CV2023

GPA-Net:No-Reference Point Cloud Quality Assessment with Multi-task Graph Convolutional Network

Ziyu Shan, Qi Yang, Rui Ye +4

With the rapid development of 3D vision, point cloud has become an increasingly popular 3D visual media content. Due to the irregular structure, point cloud has posed novel challen…

cs.MM2024

Inter-Frame Coding for Dynamic Meshes via Coarse-to-Fine Anchor Mesh Generation

He Huang, Lizhi Hou, Qi Yang +1

In the current Video-based Dynamic Mesh Coding (V-DMC) standard, inter-frame coding is restricted to mesh frames with constant topology. Consequently, temporal redundancy is not fu…

eess.IV2020

Inferring Point Cloud Quality via Graph Similarity

Qi Yang, Zhan Ma, Yiling Xu +2

We propose the GraphSIM -- an objective metric to accurately predict the subjective quality of point cloud with superimposed geometry and color impairments. Motivated by the facts…

cs.CV2026

HybridINR-PCGC: Hybrid Lossless Point Cloud Geometry Compression Bridging Pretrained Model and Implicit Neural Representation

Wenjie Huang, Qi Yang, Shuting Xia +3

Learning-based point cloud compression presents superior performance to handcrafted codecs. However, pretrained-based methods, which are based on end-to-end training and expected t…

cs.CV2025

DPCD: A Quality Assessment Database for Dynamic Point Clouds

Yating Liu, Yujie Zhang, Qi Yang +3

Recently, the advancements in Virtual/Augmented Reality (VR/AR) have driven the demand for Dynamic Point Clouds (DPC). Unlike static point clouds, DPCs are capable of capturing tem…

cs.MM2018

Viewport Adaptation-Based Immersive Video Streaming: Perceptual Modeling and Applications

Shaowei Xie, Qiu Shen, Yiling Xu +4

Immersive video offers the freedom to navigate inside virtualized environment. Instead of streaming the bulky immersive videos entirely, a viewport (also referred to as field of vi…

eess.IV2021

Which One is Better: Assessing Objective Metrics for Point Cloud Compression

Yipeng Liu, Qi Yang, Yiling Xu +1

Point cloud compression (PCC) has made remarkable achievement in recent years. In the mean time, point cloud quality assessment (PCQA) also realize gratifying development. Some rec…

cs.CV2025

3DGS-IEval-15K: A Large-scale Image Quality Evaluation Database for 3D Gaussian-Splatting

Yuke Xing, Jiarui Wang, Peizhi Niu +3

3D Gaussian Splatting (3DGS) has emerged as a promising approach for novel view synthesis, offering real-time rendering with high visual fidelity. However, its substantial storage…

eess.IV2025

Differentiable Low-computation Global Correlation Loss for Monotonicity Evaluation in Quality Assessment

Yipeng Liu, Qi Yang, Yiling Xu

In this paper, we propose a global monotonicity consistency training strategy for quality assessment, which includes a differentiable, low-computation monotonicity evaluation loss…

cs.CV2025

ADC-GS: Anchor-Driven Deformable and Compressed Gaussian Splatting for Dynamic Scene Reconstruction

He Huang, Qi Yang, Mufan Liu +2

Existing 4D Gaussian Splatting methods rely on per-Gaussian deformation from a canonical space to target frames, which overlooks redundancy among adjacent Gaussian primitives and r…

cs.CV2025

From Images to Point Clouds: An Efficient Solution for Cross-media Blind Quality Assessment without Annotated Training

Yipeng Liu, Qi Yang, Yujie Zhang +3

We present a novel quality assessment method which can predict the perceptual quality of point clouds from new scenes without available annotations by leveraging the rich prior kno…

cs.CV2024

Asynchronous Feedback Network for Perceptual Point Cloud Quality Assessment

Yujie Zhang, Qi Yang, Ziyu Shan +1

Recent years have witnessed the success of the deep learning-based technique in research of no-reference point cloud quality assessment (NR-PCQA). For a more accurate quality predi…

cs.CV2026

RAP: Fast Feedforward Rendering-Free Attribute-Guided Primitive Importance Score Prediction for Efficient 3D Gaussian Splatting Processing

Kaifa Yang, Qi Yang, Yiling Xu +1

3D Gaussian Splatting (3DGS) has emerged as a leading technology for high-quality 3D scene reconstruction. However, the iterative refinement and densification process leads to the…

cs.CV2026

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

Yuke Xing, Jiarui Wang, William Gordon +3

3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression indispensable for practical…

cs.NI2025

Video Streaming with Kairos: An MPC-Based ABR with Streaming-Aware Throughput Prediction

Ziyu Zhong, Mufan Liu, Le Yang +3

In this paper, we present Kairos, a model predictive control (MPC)-based adaptive bitrate (ABR) scheme that integrates streaming-aware throughput predictions to enhance video strea…

eess.IV2023

Reduced Reference Quality Assessment for Point Cloud Compression

Yipeng Liu, Qi Yang, Yiling Xu

In this paper, we propose a reduced reference (RR) point cloud quality assessment (PCQA) model named R-PCQA to quantify the distortions introduced by the lossy compression. Specifi…

cs.CV2022

H2-Stereo: High-Speed, High-Resolution Stereoscopic Video System

Ming Cheng, Yiling Xu, Wang Shen +4

High-speed, high-resolution stereoscopic (H2-Stereo) video allows us to perceive dynamic 3D content at fine granularity. The acquisition of H2-Stereo video, however, remains challe…

cs.CV2022

D-DPCC: Deep Dynamic Point Cloud Compression via 3D Motion Prediction

Tingyu Fan, Linyao Gao, Yiling Xu +2

The non-uniformly distributed nature of the 3D dynamic point cloud (DPC) brings significant challenges to its high-efficient inter-frame compression. This paper proposes a novel 3D…

cs.CV2024

Perception-Guided Quality Metric of 3D Point Clouds Using Hybrid Strategy

Yujie Zhang, Qi Yang, Yiling Xu +1

Full-reference point cloud quality assessment (FR-PCQA) aims to infer the quality of distorted point clouds with available references. Most of the existing FR-PCQA metrics ignore t…

cs.CV2023

TCDM: Transformational Complexity Based Distortion Metric for Perceptual Point Cloud Quality Assessment

Yujie Zhang, Qi Yang, Yifei Zhou +3

The goal of objective point cloud quality assessment (PCQA) research is to develop quantitative metrics that measure point cloud quality in a perceptually consistent manner. Mergin…

cs.MM2024

StreamOptix: A Cross-layer Adaptive Video Delivery Scheme

Mufan Liu, Le Yang, Yifan Wang +3

This paper presents a cross-layer video delivery scheme, StreamOptix, and proposes a joint optimization algorithm for video delivery that leverages the characteristics of the physi…

cs.CV2022

Point Cloud Quality Assessment using 3D Saliency Maps

Zhengyu Wang, Yujie Zhang, Qi Yang +3

Point cloud quality assessment (PCQA) has become an appealing research field in recent days. Considering the importance of saliency detection in quality assessment, we propose an e…

cs.CV2025

Standardizing Generative Face Video Compression using Supplemental Enhancement Information

Bolin Chen, Yan Ye, Jie Chen +12

This paper proposes a Generative Face Video Compression (GFVC) approach using Supplemental Enhancement Information (SEI), where a series of compact spatial and temporal representat…

cs.CV2025

LINR-PCGC: Lossless Implicit Neural Representations for Point Cloud Geometry Compression

Wenjie Huang, Qi Yang, Shuting Xia +3

Existing AI-based point cloud compression methods struggle with dependence on specific training data distributions, which limits their real-world deployment. Implicit Neural Repres…

cs.CV2025

CLIP-PCQA: Exploring Subjective-Aligned Vision-Language Modeling for Point Cloud Quality Assessment

Yating Liu, Yujie Zhang, Ziyu Shan +1

In recent years, No-Reference Point Cloud Quality Assessment (NR-PCQA) research has achieved significant progress. However, existing methods mostly seek a direct mapping function f…

cs.CV2024

PAME: Self-Supervised Masked Autoencoder for No-Reference Point Cloud Quality Assessment

Ziyu Shan, Yujie Zhang, Qi Yang +3

No-reference point cloud quality assessment (NR-PCQA) aims to automatically predict the perceptual quality of point clouds without reference, which has achieved remarkable performa…

cs.CV2025

A Hierarchical Compression Technique for 3D Gaussian Splatting Compression

He Huang, Wenjie Huang, Qi Yang +2

3D Gaussian Splatting (GS) demonstrates excellent rendering quality and generation speed in novel view synthesis. However, substantial data size poses challenges for storage and tr…

eess.IV2022

Point Cloud Quality Assessment: Dataset Construction and Learning-based No-Reference Metric

Yipeng Liu, Qi Yang, Yiling Xu +1

Full-reference (FR) point cloud quality assessment (PCQA) has achieved impressive progress in recent years. However, in many cases, obtaining the reference point clouds is difficul…

cs.CV2022

MPED: Quantifying Point Cloud Distortion based on Multiscale Potential Energy Discrepancy

Qi Yang, Yujie Zhang, Siheng Chen +3

In this paper, we propose a new distortion quantification method for point clouds, the multiscale potential energy discrepancy (MPED). Currently, there is a lack of effective disto…

cs.CV2022

No-Reference Point Cloud Quality Assessment via Domain Adaptation

Qi Yang, Yipeng Liu, Siheng Chen +2

We present a novel no-reference quality assessment metric, the image transferred point cloud quality assessment (IT-PCQA), for 3D point clouds. For quality assessment, deep neural…

cs.CV2025

Point Cloud Compression and Objective Quality Assessment: A Survey

Yiling Xu, Yujie Zhang, Shuting Xia +6

The rapid growth of 3D point cloud data, driven by applications in autonomous driving, robotics, and immersive environments, has led to criticals demand for efficient compression a…

cs.CV2024

A Benchmark for Gaussian Splatting Compression and Quality Assessment Study

Qi Yang, Kaifa Yang, Yuke Xing +2

To fill the gap of traditional GS compression method, in this paper, we first propose a simple and effective GS data compression anchor called Graph-based GS Compression (GGSC). GG…

cs.CV2026

Light4GS: Lightweight Compact 4D Gaussian Splatting Generation via Context Model

Mufan Liu, Qi Yang, He Huang +4

3D Gaussian Splatting (3DGS) has emerged as an efficient and high-fidelity paradigm for novel view synthesis. To adapt 3DGS for dynamic content, deformable 3DGS incorporates tempor…

cs.CV2023

Multiscale Latent-Guided Entropy Model for LiDAR Point Cloud Compression

Tingyu Fan, Linyao Gao, Yiling Xu +2

The non-uniform distribution and extremely sparse nature of the LiDAR point cloud (LPC) bring significant challenges to its high-efficient compression. This paper proposes a novel…

cs.CV2025

Towards Fine-Grained Text-to-3D Quality Assessment: A Benchmark and A Two-Stage Rank-Learning Metric

Bingyang Cui, Yujie Zhang, Qi Yang +2

Recent advances in Text-to-3D (T23D) generative models have enabled the synthesis of diverse, high-fidelity 3D assets from textual prompts. However, existing challenges restrict th…

cs.CV2022

4DAC: Learning Attribute Compression for Dynamic Point Clouds

Guangchi Fang, Qingyong Hu, Yiling Xu +1

With the development of the 3D data acquisition facilities, the increasing scale of acquired 3D point clouds poses a challenge to the existing data compression techniques. Although…

eess.IV2026

TVMC: Time-Varying Mesh Compression via Multi-Stage Anchor Mesh Generation

He Huang, Qi Yang, Yiling Xu +2

Time-varying meshes, characterized by dynamic connectivity and varying vertex counts, hold significant promise for applications such as augmented reality. However, their practical…

cs.CV2026

Progressively Deformable 2D Gaussian Splatting for Video Representation at Arbitrary Resolutions

Mufan Liu, Qi Yang, Miaoran Zhao +4

Implicit neural representations (INRs) enable fast video compression and effective video processing, but a single model rarely offers scalable decoding across rates and resolutions…

eess.IV2019

Learned Quality Enhancement via Multi-Frame Priors for HEVC Compliant Low-Delay Applications

Ming Lu, Ming Cheng, Yiling Xu +3

Networked video applications, e.g., video conferencing, often suffer from poor visual quality due to unexpected network fluctuation and limited bandwidth. In this paper, we have de…

cs.CV2024

Learning Disentangled Representations for Perceptual Point Cloud Quality Assessment via Mutual Information Minimization

Ziyu Shan, Yujie Zhang, Yipeng Liu +1

No-Reference Point Cloud Quality Assessment (NR-PCQA) aims to objectively assess the human perceptual quality of point clouds without relying on pristine-quality point clouds for r…

cs.CV2024

Contrastive Pre-Training with Multi-View Fusion for No-Reference Point Cloud Quality Assessment

Ziyu Shan, Yujie Zhang, Qi Yang +5

No-reference point cloud quality assessment (NR-PCQA) aims to automatically evaluate the perceptual quality of distorted point clouds without available reference, which have achiev…

eess.IV2020

A Dual Camera System for High Spatiotemporal Resolution Video Acquisition

Ming Cheng, Zhan Ma, M. Salman Asif +4

This paper presents a dual camera system for high spatiotemporal resolution (HSTR) video acquisition, where one camera shoots a video with high spatial resolution and low frame rat…