A Non-Autoregressive Text-to-Speech (NAR-TTS) framework, including official PyTorch implementation of PortaSpeech (NeurIPS 2021) and DiffSpeech (AAAI 2022)
-
Updated
Apr 2, 2023 - Python
A Non-Autoregressive Text-to-Speech (NAR-TTS) framework, including official PyTorch implementation of PortaSpeech (NeurIPS 2021) and DiffSpeech (AAAI 2022)
PyTorch implementation of DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (focused on DiffSpeech)
Singing Voice Synthesis based on VITS, different from VISinger
A python GUI toolkit for creating/editing Aesthetic YAML dictionaries for OpenUtau
Multispeaker Community Vocoder Model for DiffSinger
🎼🎵𝐄𝐱𝐩𝐫𝐞𝐬𝐬𝐢𝐯𝐞 | 适用于OpenUtau的DiffSinger歌手表情参数导入工具。从真实歌手的人声中提取表情,并导入到工程的相应轨道上 Migrate expressions from real singers to DiffSingers
Convert the UTAU Voicebank to a configuration compatible with DiffSinger Dataset
The open-source alternative to Suno and ElevenLabs Music. Natural language music composition, run locally, own everything.
A fork of genon2nnsvs with modifications made for english speakers and diffsinger users
a CustomTkInter GUI for processing and training DiffSinger models
Thai RVC WebUI is edit and use for develop to our Voicebank for the best Quality.
Deep learning-based singing voice synthesis project using DiffSinger with phoneme alignment, Mel-spectrogram and F0 feature extraction.
Real-time DiffSinger synthesis on x86-64 CPUs with C and handwritten assembly
Singing Voice Synthesis via Shallow Diffusion Mechanism: explore phoneme-mapped cross-lingual transfer learning using minimal target language data (English to German)
Step-by-step toolkit for DiffSinger voice synthesis. Preprocessing scripts + configs + guides to go from raw audio to trained singing voice models.
i proved that a word to phoneme system Is possible.
To associate your repository with the diffsinger topic, visit your repo's landing page and select "manage topics."