activity
20242026
collaborators

5 papers

cs.RO2026

: Unified Chain of Perception-Prediction-Planning Thought via Reinforcement Fine-Tuning

Yuqi Ye, Zijian Zhang, Junhong Lin +3

Vision-language models (VLMs) are increasingly being adopted for end-to-end autonomous driving systems due to their exceptional performance in handling long-tail scenarios. However…

cs.CV2025

AnyPcc: Compressing Any Point Cloud with a Single Universal Model

Kangli Wang, Qianxi Yi, Yuqi Ye +2

Generalization remains a critical challenge in deep learning-based point cloud geometry compression. While existing methods perform well on standard benchmarks, their performance c…

cs.CV2025

Generalized Gaussian Entropy Model for Point Cloud Attribute Compression with Dynamic Likelihood Intervals

Changhao Peng, Yuqi Ye, Wei Gao

Gaussian and Laplacian entropy models are proved effective in learned point cloud attribute compression, as they assist in arithmetic coding of latents. However, we demonstrate thr…

cs.SD2025

STFTCodec: High-Fidelity Audio Compression through Time-Frequency Domain Representation

Tao Feng, Zhiyuan Zhao, Yifan Xie +4

We present STFTCodec, a novel spectral-based neural audio codec that efficiently compresses audio using Short-Time Fourier Transform (STFT). Unlike waveform-based approaches that r…

cs.AI2024

LLM-PCGC: Large Language Model-based Point Cloud Geometry Compression

Yuqi Ye, Wei Gao

The key to effective point cloud compression is to obtain a robust context model consistent with complex 3D data structures. Recently, the advancement of large language models (LLM…