2 papers
cs.CV2026
ELDiff: When Evidential Learning Meets Text-to-Image Diffusion
Qingtao Pan, Kai Ye, Zhihao Dou +2
In multi-object text-to-image (T2I) diffusion, ensuring semantic consistency between textual prompts and generated visual content is crucial for image synthesis. However, such cons…
cs.CV2024
HMANet: Hybrid Multi-Axis Aggregation Network for Image Super-Resolution
Shu-Chuan Chu, Zhi-Chao Dou, Jeng-Shyang Pan +2
Transformer-based methods have demonstrated excellent performance on super-resolution visual tasks, surpassing conventional convolutional neural networks. However, existing work ty…