Basic Pitch Ts
spotify
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 38 · 16 shown · 743 public entries
Results: 743
spotify
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection.
Renovamen
Speech to text (PocketSphinx, Iflytex API, Baidu API) and text to speech (pyttsx3) | 语音转文字(PocketSphinx、百度 API、科大讯飞 API)和文字转语音(pyttsx3)
henkisdabro
FFmpeg CLI reference for video and audio processing, format conversion, filtering, and media automation. Use when converting video formats, resizing or cropping video, t…
AsyrafHussin
Laravel AI SDK for building AI-powered features. Use when creating agents, generating images or audio, working with embeddings, vector search, or testing AI features. Tr…
Rongjiehuang
PyTorch Implementation of GenerSpeech (NeurIPS'22): a text-to-speech model towards zero-shot style transfer of OOD custom voice.
keonlee9420
A Non-Autoregressive Transformer based Text-to-Speech, supporting a family of SOTA transformers with supervised and unsupervised duration modelings. This project grows w…
seanwood
Real-time GCC-NMF Blind Speech Separation and Enhancement
rishikksh20
VocGAN: A High-Fidelity Real-time Vocoder with a Hierarchically-nested Adversarial Network
keonlee9420
PyTorch Implementation of Non-autoregressive Expressive (emotional, conversational) TTS based on FastSpeech2, supporting English, Korean, and your own languages.
This is a python application which converts american sign language into text and speech which helps Dumb/Deaf people to start conversation with normal people who dont un…
This AI Smart Speaker uses speech recognition, TTS (text-to-speech), and STT (speech-to-text) to enable voice and vision-driven conversations, with additional web search…
AlterLab-IEU
Infer gene regulatory networks (GRNs) from expression matrices using arboreto's scalable GRNBoost2 and GENIE3 tree-ensemble algorithms with Dask-distributed computation.…
roedyrustam
Expert guide for AI image generation (Flux, DALL-E, Stable Diffusion), video generation (Sora, Runway), voice synthesis (ElevenLabs TTS), and speech recognition (Whisper…
mbzuai-oryx
LLMVoX: Autoregressive Streaming Text-to-Speech Model for Any LLM
dhvcc
Parse RSS 2.0/0.9x, Atom 1.0, RSS 1.0 (RDF) and podcast (itunes:*) feeds in Python into typed pydantic v2 models with the `rss-parser` package. Use this skill whenever a…
dodiku
Fast and simple music and audio analysis using RNN in Python 🕵️♀️ 🥁