2 papers
cs.GR2026
VectorGym: A Multi-Task Benchmark for SVG Code Generation, Sketching and Editing
Joan Rodriguez, Haotian Zhang, Abhay Puri +14
We introduce VectorGym, a comprehensive benchmark suite for Scalable Vector Graphics (SVG) that spans generation from text and sketches, complex editing, and visual understanding.…
cs.CV2016
Areas of Attention for Image Captioning
Marco Pedersoli, Thomas Lucas, Cordelia Schmid +1
We propose "Areas of Attention", a novel attention-based model for automatic image captioning. Our approach models the dependencies between image regions, caption words, and the st…