No visual example yet
Explore the skillExpo Speech Recognition
jamsch
Speech Recognition for React Native Expo projects
OPENAGENTSKILL / DIRECTORY
Finde den passenden Skill für deine nächste Aufgabe mit Codex, Claude Code, Cursor und mehr.
90 Skills
Ergebnisse: 90
No visual example yet
Explore the skilljamsch
Speech Recognition for React Native Expo projects
No visual example yet
Explore the skillhuggingface
The official Python client for the Hugging Face Hub.
No visual example yet
Explore the skillhuggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
No visual example yet
Explore the skillcalesthio
Swap faces in a video using AI via the HeyGen API. Use when: (1) Replacing a face in a video with another face, (2) Face swapping from a source image onto a target video…
No visual example yet
Explore the skillPaddlePaddle
Use this skill whenever the user wants text extracted from images, photos, scans, screenshots, or scanned PDFs. Returns exact machine-readable strings with line
No visual example yet
Explore the skillyeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…
No visual example yet
Explore the skillm-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
No visual example yet
Explore the skillflashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
No visual example yet
Explore the skillFunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
No visual example yet
Explore the skillalphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
No visual example yet
Explore the skillCMU-Perceptual-Computing-Lab
OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation
No visual example yet
Explore the skillsemperai
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
No visual example yet
Explore the skillTalAter
💬 Speech recognition for your site
No visual example yet
Explore the skillmindee
docTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks powered by Deep Learning.
No visual example yet
Explore the skillwenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
No visual example yet
Explore the skillMahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper