Basic Pitch
spotify
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Search retrieves candidates across the registry and ranks a bounded shortlist by task fit. This count is matching candidates, not the registry total. No suitable match? Try a specific tool or task.
17β32 / 63
Results: 63
spotify
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection
rsxdalv
A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Auβ¦
milvus-io
Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video analysis, question and answer systems, NLP, etc.
bytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
pluja
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!
enhuiz
An unofficial PyTorch implementation of the audio LM VALL-E
readbeyond
aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment)
DamRsn
Audio Plugin for Audio to MIDI transcription using deep learning.
Pluviobyte
Generate a controlled local narration workflow with auditions, version tracking, and subtitle-ready final audio.
hugohe3
AI generates a real, editable PowerPoint from any document β native shapes & animations, speaker notes voiced as audio narration, and the option to follow your own .pptxβ¦
huggingface
π€ Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
myshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
Eventual-Inc
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
clovaai
Official Implementation of OCR-free Document Understanding Transformer (Donut) and Synthetic Document Generator (SynthDoG), ECCV 2022
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.