Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
-
Updated
Apr 24, 2024 - Python
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
The first medical SpeechLM, open-sourced with weight, data, and code of training, inference, and evaluation.
Open-source neural speech synthesis for pre-training, SFT, and RLHF on SpeechLM and audio codecs across local and cloud GPUs.
Add a description, image, and links to the speechlm topic page so that developers can more easily learn about it.
To associate your repository with the speechlm topic, visit your repo's landing page and select "manage topics."