5 papers
MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning
Zhihui Chen, Kai He, Qingyuan Lei +4
Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinical trust and safety. Existing def…
Med-Banana: Learning Quality-Controlled Medical Image Editing from Success-and-Failure Trajectories
Zhihui Chen, Qingyuan Lei, Kai He +2
Text-guided medical image editing must satisfy the requested pathology while preserving anatomy, modality-specific appearance, and clinical plausibility. However, existing datasets…
HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer
Qi Cai, Jingwen Chen, Chengmin Gao +22
The evolution of visual generative models has long been constrained by fragmented architectures relying on disjoint text encoders and external VAEs. In this report, we present HiDr…
medR: Reward Engineering for Clinical Offline Reinforcement Learning via Tri-Drive Potential Functions
Qianyi Xu, Gousia Habib, Feng Wu +5
Reinforcement Learning (RL) offers a powerful framework for optimizing dynamic treatment regimes (DTRs). However, clinical RL is fundamentally bottlenecked by reward engineering: t…
DivScore: Zero-Shot Detection of LLM-Generated Text in Specialized Domains
Zhihui Chen, Kai He, Yucheng Huang +2
Detecting LLM-generated text in specialized and high-stakes domains like medicine and law is crucial for combating misinformation and ensuring authenticity. However, current zero-s…