1 paper
Abhineet Singh, Justin Rozeboom, Nilanjan Ray
This paper presents a new unified approach to semantic segmentation in both images and videos by using language modeling to output the masks as sequences of discrete tokens. We use…