13 papers
Universal Image Restoration via Internalized Chain-of-Thought Reasoning
Yu Guo, Zhengru Fang, Shengfeng He +4
Image restoration seeks to recover high-quality images from degraded inputs but becomes highly ill-posed under complex, mixed degradations. While unified all-in-one models are comm…
Optimizing Agentic Reasoning with Retrieval via Synthetic Semantic Information Gain Reward
Senkang Hu, Yong Dai, Yuzhi Zhao +5
Agentic reasoning enables large reasoning models (LRMs) to dynamically acquire external knowledge, but yet optimizing the retrieval process remains challenging due to the lack of d…
FRUC: Feedforward Dynamic Scene Reconstruction from Uncalibrated Collaborative Driving Views
Yihang Tao, Yu Guo, Zhengru Fang +2
We present FRUC, a feed-forward 3D Gaussian splatting framework for dynamic scene reconstruction from uncalibrated collaborative driving views. Existing multi-agent reconstruction…
V2VCrafter: Consistent Street-View Image Generation Across Vehicles
Yihang Tao, Yu Guo, Senkang Hu +4
Connected and autonomous driving (CAD) systems leverage vehicle-to-vehicle (V2V) communication for multi-agent collaborative perception, yet remain constrained by scarce annotated…
Agent-Centric Observation Adaptation for Robust Visual Control under Dynamic Perturbations
Zhengru Fang, Yu Guo, Fei Liu +5
Real-world visual systems face time-varying perturbations, including weather, sensor noise, compression artifacts, and background distractions. Existing image restoration methods a…
CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs
Zhengru Fang, Yanan Ma, Yu Guo +5
When a chest X-ray shows consolidation but the question asks which finding is present, a medical vision-language model may answer "No consolidation." This is more than an incorrect…