activity
20242026
most citedBeyond Gaussians: Fast and High-Fidelity 3D Splatting with Linear Kernels

3 citations · 3 across the 23 of their papers we have counts for

collaborators

28 papers

cs.CV2026

Delving into the Temporal Challenges of Unified Video Protection Against Image-to-Video and Fine-Tuning-based Customization

Yuxin Huang, Ziming Hong, Mingming Gong +3

Recent diffusion-based video generation models have enabled high-quality personalized video customization through both tuning-based pipelines, which fine-tune a video diffusion mod…

cs.LG2026

Exploring the Design Space of Reward Backpropagation for Flow Matching

Ruoyu Wang, Boye Niu, Xiangxin Zhou +3

Aligning text-to-image flow matching models with human preferences via direct reward backpropagation is sample-efficient but hampered by two well-known pathologies: activations can…

cs.LG2026

Causal Ensemble Agent: Hierarchical Causal Discovery with LLM-guided Expert Reweighting

Xinyu Li, Yuanyuan Wang, Haoxuan Li +7

Causal discovery aims to uncover causal structures from observational data, which is crucial for real-world decision-making. However, different causal discovery algorithms can prod…

cs.GR2026

AnisoLift: Anisotropic Latent Representations for Coarse Particle Liquid Enhancement

Zhengqing Gao, Huaxi Huang, Runqi Lin +6

Particle-based liquid simulation is widely used in graphics and physical modeling, but high-resolution rollouts remain computationally expensive. Consequently, many methods aim to…

cs.CV2026

LightHarmony3D: Harmonizing Illumination and Shadows for Object Insertion in 3D Gaussian Splatting

Tianyu Huang, Zhenyang Ren, Zhenchen Wan +5

3D Gaussian Splatting (3DGS) enables high-fidelity reconstruction of scene geometry and appearance. Building on this capability, inserting external mesh objects into reconstructed…

cs.RO2026

KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition

Gaoge Han, Zhengqing Gao, Ziwen Li +5

In this paper, we introduce a novel kinematics-rich vision-language-action (VLA) task, in which language commands densely encode diverse kinematic attributes (such as direction, tr…