most citedMediAug: Exploring Visual Augmentation in Medical Imaging

1 citations · 2 across the 7 of their papers we have counts for

collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2025

PresentAgent: Multimodal Agent for Presentation Video Generation

Jingwei Shi, Zeyu Zhang, Biao Wu +4

We present PresentAgent, a multimodal agent that transforms long-form documents into narrated presentation videos. While existing approaches are limited to generating static slides…

cs.CV2025

DC-Scene: Data-Centric Learning for 3D Scene Understanding

Ting Huang, Zeyu Zhang, Ruicheng Zhang +1

3D scene understanding plays a fundamental role in vision applications such as robotics, autonomous driving, and augmented reality. However, advancing learning-based 3D scene under…

cs.CV2025★ 1 cited

MediAug: Exploring Visual Augmentation in Medical Imaging

Xuyin Qi, Zeyu Zhang, Canxuan Gang +4

Data augmentation is essential in medical imaging for improving classification accuracy, lesion detection, and organ segmentation under limited data conditions. However, two signif…

cs.CV2025★ 1 cited

MedConv: Convolutions Beat Transformers on Long-Tailed Bone Density Prediction

Xuyin Qi, Zeyu Zhang, Huazhan Zheng +19

Bone density prediction via CT scans to estimate T-scores is crucial, providing a more precise assessment of bone health compared to traditional methods like X-ray bone density tes…

cs.CV2025

PedDet: Adaptive Spectral Optimization for Multimodal Pedestrian Detection

Rui Zhao, Zeyu Zhang, Yi Xu +6

Pedestrian detection in intelligent transportation systems has made significant progress but faces two critical challenges: (1) insufficient fusion of complementary information bet…

cs.CV2024

Motion Avatar: Generate Human and Animal Avatars with Arbitrary Motion

Zeyu Zhang, Yiran Wang, Biao Wu +7

In recent years, there has been significant interest in creating 3D avatars and motions, driven by their diverse applications in areas like film-making, video games, AR/VR, and hum…