activity
20232026
most citedInstilling Multi-round Thinking to Text-guided Image Generation

2 citations · 2 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV2026

Road Maps as Free Geometric Priors: Weather-Invariant Drone Geo-Localization with GeoFuse

Yunsong Fang, Tingyu Wang, Zhedong Zheng

Drone-view geo-localization aims to match a query drone image, often captured under adverse weather conditions (e.g., rain, snow, fog), against a gallery of geo-tagged satellite im…

cs.MM2025

EQ-TAA: Equivariant Traffic Accident Anticipation via Diffusion-Based Accident Video Synthesis

Jianwu Fang, Lei-Lei Li, Zhedong Zheng +4

Traffic Accident Anticipation (TAA) in traffic scenes is a challenging problem for achieving zero fatalities in the future. Current approaches typically treat TAA as a supervised l…

cs.RO2024

3D-TAFS: A Training-free Framework for 3D Affordance Segmentation

Meng Chu, Xuan Zhang, Zhedong Zheng +1

Translating high-level linguistic instructions into precise robotic actions in the physical world remains challenging, particularly when considering the feasibility of interacting…

cs.CV2024★ 2 cited

Instilling Multi-round Thinking to Text-guided Image Generation

Lidong Zeng, Zhedong Zheng, Yinwei Wei +1

This paper delves into the text-guided image editing task, focusing on modifying a reference image according to user-specified textual feedback to embody specific attributes. Despi…

cs.CV2023

Towards Natural Language-Guided Drones: GeoText-1652 Benchmark with Spatial Relation Matching

Meng Chu, Zhedong Zheng, Wei Ji +2

Navigating drones through natural language commands remains challenging due to the dearth of accessible multi-modal datasets and the stringent precision requirements for aligning v…