Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
APRIL-MedSeg: A Modular Medical Image Segmentation Toolbox Embracing Modern Paradigms
Juntao Jiang, Jinsheng Bai, Linxuan Fan +3
We present APRIL-MedSeg, a YAML-driven modular framework for 2D medical image segmentation. It provides a unified and extensible ecosystem that decomposes segmentation networks int…
cs.CV2026
JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation
Yinan Chen, Chuming Lin, Zhennan Chen +12
While instruction-based video editing has seen significant progress, joint audio-visual editing remains constrained by the absence of dedicated datasets and benchmarks. To bridge t…
cs.CV2025
Image Inversion: A Survey from GANs to Diffusion and Beyond
Yinan Chen, Jiangning Zhang, Yali Bi +6
Image inversion is a fundamental task in generative models, aiming to map images back to their latent representations to enable downstream applications such as editing, restoration…