Robuta

https://deepai.org/publication/show-and-speak-directly-synthesize-spoken-description-of-images Show and Speak: Directly Synthesize Spoken Description of Images | DeepAI Oct 23, 2020 - 10/23/20 - This paper proposes a new model, referred to as the show and speak (SAS) model that, for the first time, is able to directly synth... show andspeak directlyspoken descriptionof imagessynthesize https://aclanthology.org/W18-3910/ Varying image description tasks: spoken versus written descriptions - ACL Anthology Emiel van Miltenburg, Ruud Koolen, Emiel Krahmer. Proceedings of the Fifth Workshop on NLP for Similar Languages, Varieties and Dialects (VarDial 2018). 2018. image descriptionwritten descriptionsvaryingtasksspoken