2 papers
cs.CV2026
Can Text-to-Image Models Draw from the Right Frame of Reference?
Zheyuan Gu, Ruihang Li, Yong Huang +5
Spatial instruction following has become a crucial requirement for text-to-image (T2I) generation. A common challenge arises when directional expressions are interpreted under diff…
cs.CV2025
Learning Attribute-aware Representations for Few-shot Scene Text Segmentation
Yifan Tang, Chenming Li, Chengxu Liu +8
Supervised scene text segmentation has achieved notable progress in recent years. However, its development is largely constrained by the scarcity of high-quality datasets and the h…