venice-audio-transcription
veniceai
Transcribe audio files to text via POST /audio/transcriptions. Covers supported models (Parakeet, Whisper, Wizper, Scribe, xAI STT), supported formats (wav/flac/m4a/aac/…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 37 · 16 shown · 1,117 public entries
Results: 1117
veniceai
Transcribe audio files to text via POST /audio/transcriptions. Covers supported models (Parakeet, Whisper, Wizper, Scribe, xAI STT), supported formats (wav/flac/m4a/aac/…
veniceai
Generate and transcribe videos via Venice. Covers the async /video/quote + /video/queue + /video/retrieve + /video/complete loop, text-to-video, image-to-video, video-to…
flymin
[ICCV 2025] Official implementation of the paper “MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control”
cvondrick
Generating Videos with Scene Dynamics. NIPS 2016.
erduo1998-cell
Read and verify legacy Remotion Integrator records for a pre-v0.9 task. Never dispatch this stage in a new v1 production; the Parent assembles verified Builder media dir…
erduo1998-cell
Read and verify legacy Remotion Render records for a pre-v0.9 task. Never dispatch this stage in a new v1 production; the Parent runs deterministic preview and delivery…
wenhaochai
[CVPR 2024] MovieChat: From Dense Token to Sparse Memory for Long Video Understanding
ahmetgunduz
Real-time Hand Gesture Recognition with PyTorch on EgoGesture, NvGesture, Jester, Kinetics and UCF101
KlingAIResearch
[ICLR'25] SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints
DmitryRyumin
INTERSPEECH 2023-2024 Papers: A complete collection of influential and exciting research papers from the INTERSPEECH 2023-24 conference. Explore the latest advances in s…
thuml
Official repository for "iVideoGPT: Interactive VideoGPTs are Scalable World Models" (NeurIPS 2024), https://arxiv.org/abs/2405.15223
Jakobovski
A free audio dataset of spoken digits. An audio version of MNIST.
MasayukiSuda
This library apply video filter on generate an Mp4 and on ExoPlayer video and Video Recording with Camera2.
Jahrome907
Create and debug Minecraft 26.x and 1.21.x resource packs, including pack metadata, textures, models, blockstates, item definitions, sounds, fonts, animations, and shade…
kwsong0113
[ICML 2025] Official PyTorch Implementation of "History-Guided Video Diffusion"
CesiumGS
CesiumJS time, properties, and animation - Clock, JulianDate, TimeInterval, Property, SampledProperty, CallbackProperty, PathMode, interval and sampled path materials, i…