Matcha TTS
shivammehta25
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
33–48 / 50
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 50
shivammehta25
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
lhotse-speech
Tools for handling multimodal data in machine learning projects.
DigitalPhonetics
Controllable and fast Text-to-Speech for over 7000 languages!
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
coqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
FireRedTeam
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity gene…
voice-cloning-app
A Python/Pytorch app for easily synthesising human voices
lucidrains
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
CSTR-Edinburgh
This is now the official location of the Merlin project.
snap-research
Code for Motion Representations for Articulated Animation paper
cure-lab
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
shadow2496
Official PyTorch implementation of "VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware Normalization" (CVPR 2021)
zzh-tech
[TPAMI][ECCV2024 Oral] Clearer anytime frame interpolation & Manipulated interpolation of anything
jolibrain
Generative AI Image and Video Toolset with GANs and Diffusion for Real-World Applications
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
Picovoice
On-device wake word detection powered by deep learning