2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2026
DriveReferee: Geometric Safety Verdicts Need Not Be Learned for Driving World-Action Models
Fengcheng Yu, Dhruv Parikh, Junjie Ye +5
Generative world-action models (WAMs) jointly generate future video and vehicle actions, while their action branches remain primarily optimized by expert imitation. Yet imitation p…
cs.CV2025
Deflickering Vision-Based Occupancy Networks through Lightweight Spatio-Temporal Correlation
Fengcheng Yu, Haoran Xu, Canming Xia +2
Vision-based occupancy networks (VONs) provide an end-to-end solution for reconstructing 3D environments in autonomous driving. However, existing methods often suffer from temporal…
cs.CV2024★ 2 cited
HouseTune: Two-Stage Floorplan Generation with LLM Assistance
Ziyang Zong, Guanying Chen, Zhaohuan Zhan +2
This paper proposes a two-stage text-to-floorplan generation framework that combines the reasoning capability of Large Language Models (LLMs) with the generative power of diffusion…