Mediapipe
google-ai-edge
Cross-platform, customizable ML solutions for live and streaming media.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 39
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 39
google-ai-edge
Cross-platform, customizable ML solutions for live and streaming media.
kedro-org
Kedro is a toolbox for production-ready data science. It uses software engineering best practices to help you create data engineering and data science pipelines that are…
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
apache
Empowering Data Intelligence with Distributed SQL for Sharding, Scalability, and Security Across All Databases.
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
unslothai
Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
AaronFeng753
Video, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super Resolution VSR, SRM…
alphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Imbad0202
Academic Research Skills for Claude Code: research → write → review → revise → finalize
lightgbm-org
A fast, distributed, high performance gradient boosting (GBT, GBDT, GBRT, GBM or MART) framework based on decision tree algorithms, used for ranking, classification and…
pytorch
Datasets, Transforms and Models specific to Computer Vision
mvanhorn
Agent-led recent-trends research across social, prediction markets, video, code, and the web.
apache
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
tensorflow
An Open Source Machine Learning Framework for Everyone
huggingface
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and traini…