most citedAdaptive Language-Aware Image Reflection Removal Network

1 citations · 1 across the 4 of their papers we have counts for

collaborators

9 papers

cs.CV2026

High-Quality Exposure Correction with Diffusion-Based Image Generation Priors

Ziwen Li, Meng Cao, Jinpu Zhang +4

Although most existing exposure correction methods achieve high fidelity, they often place excessive focus on overall pixel-wise accuracy, making it challenging to effectively mode…

cs.CV2026

HUGE-Bench: A Benchmark for High-Level UAV Vision-Language-Action Tasks

Jingyu Guo, Ziye Chen, Ziwen Li +7

Existing UAV vision-language navigation (VLN) benchmarks have enabled language-guided flight, but they largely focus on long, step-wise route descriptions with goal-centric evaluat…

cs.RO2026

KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition

Gaoge Han, Zhengqing Gao, Ziwen Li +5

In this paper, we introduce a novel kinematics-rich vision-language-action (VLA) task, in which language commands densely encode diverse kinematic attributes (such as direction, tr…

cs.CV20261 cited

Adaptive Language-Aware Image Reflection Removal Network

Siyan Fang, Yuntao Wang, Jinpu Zhang +2

Existing image reflection removal methods struggle to handle complex reflections. Accurate language descriptions can help the model understand the image content to remove complex r…

cs.CV2026

Mirage2Matter: A Physically Grounded Gaussian World Model from Video

Zhengqing Gao, Ziwen Li, Xin Wang +12

The scalability of embodied intelligence is fundamentally constrained by the scarcity of real-world interaction data. While simulation platforms provide a promising alternative, ex…

cs.CV2025

PosA-VLA: Enhancing Action Generation via Pose-Conditioned Anchor Attention

Ziwen Li, Xin Wang, Hanlue Zhang +8

The Vision-Language-Action (VLA) models have demonstrated remarkable performance on embodied tasks and shown promising potential for real-world applications. However, current VLAs…