Robuta

https://deepai.org/publication/voxtlm-unified-decoder-only-models-for-consolidating-speech-recognition-synthesis-and-speech-text-continuation-tasks Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text... Sep 14, 2023 - 09/14/23 - We propose a decoder-only language model, VoxtLM, that can perform four tasks: speech recognition, speech synthesis, text generati... speech recognition synthesisdecoder only https://www.preprints.org/manuscript/202507.2282 Speech Recognition and Synthesis Models and Platforms for the Kazakh Language[v1] | Preprints.org With the rapid development of artificial intelligence and machine learning technologies, automatic speech recognition (ASR) and Text-to-Speech (TTS) are...