1 paper
Tarun Khajuria, Braian Olmiro Dias, Marharyta Domnich +1
In this work, we interpret the representations of multi-object scenes in vision encoders through the lens of structured representations. Structured representations allow modeling o…