No visual example yet
Explore the skillText2Video Zero
Picsart-AI-Research
[ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
OPENAGENTSKILL / DIRECTORY
다음 작업에 맞는 스킬을 찾아보세요. Codex, Claude Code, Cursor 등을 지원합니다.
224 Skills
검색 결과: 224
No visual example yet
Explore the skillPicsart-AI-Research
[ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
No visual example yet
Explore the skillhuggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
No visual example yet
Explore the skillkenshohara
No visual example yet
Explore the skilljustmarkham
No visual example yet
Explore the skillSandAI-org
No visual example yet
Explore the skillaudeering
No visual example yet
Explore the skillStonewuu
No visual example yet
Explore the skilljy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
No visual example yet
Explore the skillbytedance
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
No visual example yet
Explore the skillchenyme
这是一个全自动(音频)视频翻译项目。利用Whisper识别声音,AI大模型翻译字幕,最后合并字幕视频,生成翻译后的视频。
No visual example yet
Explore the skillDoubiiu
[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
No visual example yet
Explore the skillwilliamyang1991
[SIGGRAPH Asia 2023] Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation
No visual example yet
Explore the skillenhuiz
An unofficial PyTorch implementation of the audio LM VALL-E
No visual example yet
Explore the skillEvolvingLMMs-Lab
A simple, unified multimodal models training engine. Lean, flexible, and built for hacking at scale.
No visual example yet
Explore the skillDamRsn
Audio Plugin for Audio to MIDI transcription using deep learning.
No visual example yet
Explore the skillscikit-video