12 citations · 19 across the 10 of their papers we have counts for
1 paper · 1 filter
Raz Lapid, Moshe Sipper
Modern image-to-text systems typically adopt the encoder-decoder framework, which comprises two main components: an image encoder, responsible for extracting image features, and a…