Video Diffusion Pytorch
lucidrains
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
97โ112 / 169
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 169
lucidrains
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
CSTR-Edinburgh
This is now the official location of the Merlin project.
jishengpeng
[ICLR 2025] SOTA discrete acoustic codec models with 40/75 tokens per second for audio language modeling
sdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within yourโฆ
gitmylo
A webui for different audio related Neural Networks
zai-org
CogView4, CogView3-Plus and CogView3(ECCV 2024)
savbell
๐ฌ๐ A small dictation app using OpenAI's Whisper speech recognition model.
omerbt
Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" (ICML 2023)
showlab
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
alesaccoia
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
ManimCommunity
Manim plugin for all things voiceover
Vonage
Vonage REST API client for PHP. API support for SMS, Voice, Text-to-Speech, Numbers, Verify (2FA) and more.
sandrohanea
Whisper.net. Speech to text made simple using Whisper Models
TheStageAI
Optimized Whisper models for streaming and on-device use