Skill ディレクトリ

AI Agent のための再利用可能な Skill を見つける。

タスクで実際の GitHub Skill を検索し、利用前に Stars、Trust、監査、カテゴリ、インストール経路を確認できます。

すべての推奨は、リポジトリ、監査、インストール経路に明確につながっています。

検索結果: media

英語版ディレクトリ

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

23K
Stars
87/100
信頼
カテゴリ: media-automation監査

Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.

67K
Stars
87/100
信頼
カテゴリ: media-automation監査

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

59K
Stars
87/100
信頼
カテゴリ: media-automation監査

Port of OpenAI's Whisper model in C/C++

51K
Stars
82/100
信頼
カテゴリ: media-automation監査

Video, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super Resolution VSR, SRMD, RealSR, Anime4K, RIFE, IFRNet, CAIN, DAIN, and ACNet.

17K
Stars
80/100
信頼
カテゴリ: media-automation監査

Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

15K
Stars
87/100
信頼
カテゴリ: media-automation監査

Cross-platform, customizable ML solutions for live and streaming media.

36K
Stars
82/100
信頼
カテゴリ: ml-automation監査

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

34K
Stars
87/100
信頼
カテゴリ: media-automation監査

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

31K
Stars
87/100
信頼
カテゴリ: media-automation監査

🔒 Consolidating and extending hosts files from several well-curated sources. Optionally pick extensions for porn, social media, and other categories.

31K
Stars
87/100
信頼
カテゴリ: legal-compliance監査

Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.

27K
Stars
87/100
信頼
カテゴリ: media-automation監査

Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.

8.6K
Stars
77/100
信頼
カテゴリ: media-automation監査