2 papers
cs.CV2024
Resolving Inconsistent Semantics in Multi-Dataset Image Segmentation
Qilong Zhangli, Di Liu, Abhishek Aich +2
Leveraging multiple training datasets to scale up image segmentation models is beneficial for increasing robustness and semantic understanding. Individual datasets have well-define…
cs.CV2024
Layout Agnostic Scene Text Image Synthesis with Diffusion Models
Qilong Zhangli, Jindong Jiang, Di Liu +6
While diffusion models have significantly advanced the quality of image generation their capability to accurately and coherently render text within these images remains a substanti…