Tensor Stream
osai-ai
A library for real-time video stream decoding to CUDA memory
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Browse all public registry entries, one page at a time. Counts include resources; MCP-only resources are omitted from displayed skills. Listing is not a safety or compatibility guarantee.
Page 44 · 16 shown · 824 public entries
Results: 824
osai-ai
A library for real-time video stream decoding to CUDA memory
SimformSolutionsPvtLtd
This is a library of FFmpeg for android... 📸 🎞 🚑
aws-samples
A working prototype for capturing frames off of a live MJPEG video stream, identifying objects in near real-time using deep learning, and triggering actions based on an…
X-LANCE
[ICASSP 2024] This is the official code for "VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching"
zhangmozhe
The source code of CVPR 2019 paper "Deep Exemplar-based Video Colorization".
NickLucche
GPU-ready Dockerfile to run Stability.AI stable-diffusion model v2 with a simple web interface. Includes multi-GPUs support.
scopeInfinity
Video to Text: Natural language description generator for some given video. [Video Captioning]
Nikorasu
A nearly-live implementation of OpenAI's Whisper, using sounddevice. Requires existing Whisper install.
deterministic-algorithms-lab
Tacotron 2 - PyTorch implementation with faster-than-realtime inference modified to enable cross lingual voice cloning.
FotographerAI
In-context subject-driven image generation while preserving foreground fidelity
megaease
EaseVoice Trainer is a simple and user-friendly voice cloning and speech model trainer.
EtienneAb3d
Experimental code: sound file preprocessing to optimize Whisper transcriptions without hallucinated texts
keonlee9420
PyTorch Implementation of DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs
rhulha
Unlimited text-to-speech in the Browser using Kokoro-JS, 100% local, 100% open source
keonlee9420
PyTorch Implementation of PortaSpeech: Portable and High-Quality Generative Text-to-Speech
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
RedSkill · 流白Livo · 1.0.0
Describe a rain curtain, growing flowering branches or a flock of swallows. This RedSkill package adapts three p5.js templates into sketch.js code with custom colors, density and speed.
Noncommercial use only · Local runtime recording · Use on the source platform