PyTorch Implementation of Non-autoregressive Expressive (emotional, conversational) TTS based on FastSpeech2, supporting English, Korean, and your own languages.
$ npx skills add keonlee9420/Expressive-FastSpeech2Scenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets