2 citations · 2 across the 6 of their papers we have counts for
1 paper · 1 filter
Dasen Dai, Shuoqi Li, Ronghao Chen +3
UI-to-Code generation requires vision-language models (VLMs) to produce thousands of tokens of structured HTML/CSS from a single screenshot, making visual token efficiency critical…