1 paper
Ajay Vikram Periasami, Junlin Wang, Bhuwan Dhingra
Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Existing benchmarks either focus…