3 papers
cs.CV2026
Embedding Rotation Invariance for Provable Multi-Oriented Scene Text Recognition
Zhibin Ma, Pengwen Dai, Yi Liu +3
Multi-oriented text is ubiquitous in real-world scenes and remains a major challenge for scene text recognition (STR). Existing rotation-aware methods explicitly estimate text orie…
cs.CV2026
Towards Training-Free Scene Text Editing
Yubo Li, Xugong Qin, Peng Zhang +3
Scene text editing seeks to modify textual content in natural images while maintaining visual realism and semantic consistency. Existing methods often require task-specific trainin…
cs.CV2024
Focus, Distinguish, and Prompt: Unleashing CLIP for Efficient and Flexible Scene Text Retrieval
Gangyan Zeng, Yuan Zhang, Jin Wei +5
Scene text retrieval aims to find all images containing the query text from an image gallery. Current efforts tend to adopt an Optical Character Recognition (OCR) pipeline, which r…