1 paper · 1 filter
Natan Bagrov, Eugene Khvedchenia, Borys Tymchenko +9
Vision-language models (VLMs) have recently expanded from static image understanding to video reasoning, but their scalability is fundamentally limited by the quadratic cost of pro…