Skill 디렉토리

AI Agent를 위한 재사용 가능한 Skill을 찾으세요.

작업으로 실제 GitHub Skill을 검색하고 사용 전에 Stars, 신뢰, 감사, 카테고리, 설치 경로를 확인하세요.

모든 추천은 리포지토리, 감사, 설치 경로와 명확하게 연결됩니다.

검색 결과: media

영문 디렉토리

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

23K
Stars
87/100
신뢰
카테고리: media-automation감사

Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.

67K
Stars
87/100
신뢰
카테고리: media-automation감사

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

59K
Stars
87/100
신뢰
카테고리: media-automation감사

Port of OpenAI's Whisper model in C/C++

51K
Stars
82/100
신뢰
카테고리: media-automation감사

Video, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super Resolution VSR, SRMD, RealSR, Anime4K, RIFE, IFRNet, CAIN, DAIN, and ACNet.

17K
Stars
80/100
신뢰
카테고리: media-automation감사

Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

15K
Stars
87/100
신뢰
카테고리: media-automation감사

Cross-platform, customizable ML solutions for live and streaming media.

36K
Stars
82/100
신뢰
카테고리: ml-automation감사

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

34K
Stars
87/100
신뢰
카테고리: media-automation감사

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

31K
Stars
87/100
신뢰
카테고리: media-automation감사

🔒 Consolidating and extending hosts files from several well-curated sources. Optionally pick extensions for porn, social media, and other categories.

31K
Stars
87/100
신뢰
카테고리: legal-compliance감사

Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.

27K
Stars
87/100
신뢰
카테고리: media-automation감사

Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.

8.6K
Stars
77/100
신뢰
카테고리: media-automation감사