2 papers
cs.CV2026
FontUse: A Data-Centric Approach to Style- and Use-Case-Conditioned In-Image Typography
Xia Xin, Yuki Endo, Yoshihiro Kanamori
Recent text-to-image models can generate high-quality images from natural-language prompts, yet controlling typography remains challenging: requested typographic appearance is ofte…
cs.SD2025
The ICME 2025 Audio Encoder Capability Challenge
Junbo Zhang, Heinrich Dinkel, Qiong Song +8
This challenge aims to evaluate the capabilities of audio encoders, especially in the context of multi-task learning and real-world applications. Participants are invited to submit…