Aeneas
readbeyond
aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment)
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
177–192 / 476
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 476
readbeyond
aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment)
TMElyralab
MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
ai-forever
Kandinsky 2 — multilingual text2image latent diffusion model
janvarev
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
lhotse-speech
Tools for handling multimodal data in machine learning projects.
byjlw
Analyze videos using LLMs, Computer Vision and Automatic Speech Recognition
coqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
marytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
ErickWendel
JS Expert Week 8.0 - 🎥Pre processing videos before uploading in the browser 😏
mravanelli
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label…
jik876
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
Rayhane-mamah
DeepMind's Tacotron-2 Tensorflow implementation