No visual example yet
Explore the skillVibeVoice ComfyUI
Enemyx-net
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your C…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
209–224 / 466
Results: 466
No visual example yet
Explore the skillEnemyx-net
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your C…
No visual example yet
Explore the skillbackblaze-labs
Genblaze is an open source Python SDK for orchestrating generative AI media pipelines across video, audio, and image providers with built in provenance for every output.
No visual example yet
Explore the skillapocas
RESTai is an AIaaS (AI as a Service) open-source platform. Supports many public and local LLM suported by Ollama/vLLM/etc. Precise embeddings usage, tuning, analytics et…
No visual example yet
Explore the skillSoftcatala
Whisper command line client compatible with original OpenAI client based on CTranslate2.
No visual example yet
Explore the skillMoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
No visual example yet
Explore the skillGauravSingh9356
Personal Assistant built using python libraries. It does almost anything which includes sending emails, Optical Text Recognition, Dynamic News Reporting at any time with…
No visual example yet
Explore the skillsanchit-gandhi
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
No visual example yet
Explore the skillFrancis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-pro…
No visual example yet
Explore the skillekwek1
Soprano: Instant, Ultra-Realistic Text-to-Speech
No visual example yet
Explore the skillTencent-Hunyuan
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
No visual example yet
Explore the skillavevlad
Список русскоязычных подкастов на тему информационных технологий
No visual example yet
Explore the skillwladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
No visual example yet
Explore the skillnari-labs
TTS model capable of streaming conversational audio in realtime.
No visual example yet
Explore the skillmetavoiceio
Foundational model for human-like, expressive TTS
No visual example yet
Explore the skillsooftware
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
No visual example yet
Explore the skillhuggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Stock analysis, market research, quant backtesting, financial data, and investment research skills.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.
Video-generation prompts, B-roll, Vox-style explainers, camera direction, captions, and creative-production workflows.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.