9 citations · 9 across the 1 of their papers we have counts for
4 papers
ZR-2021VG: Zero-Resource Speech Challenge, Visually-Grounded Language Modelling track, 2021 edition
Afra Alishahi, Grzegorz Chrupała, Alejandrina Cristia +5
We present the visually-grounded language modelling track that was introduced in the Zero-Resource Speech challenge, 2021 edition, 2nd round. We motivate the new track and discuss…
Discrete representations in neural models of spoken language
Bertrand Higy, Lieke Gelderloos, Afra Alishahi +1
The distributed and continuous representations used by neural networks are at odds with representations employed in linguistics, which are typically symbolic. Vector quantization h…
Textual Supervision for Visually Grounded Spoken Language Understanding
Bertrand Higy, Desmond Elliott, Grzegorz Chrupała
Visually-grounded models of spoken language understanding extract semantic information directly from speech, without relying on transcriptions. This is useful for low-resource lang…
Few-shot learning with attention-based sequence-to-sequence models
Bertrand Higy, Peter Bell
End-to-end approaches have recently become popular as a means of simplifying the training and deployment of speech recognition systems. However, they often require large amounts of…