Speech Recognition
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
49–64 / 108
Results: 108
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
NVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
debpalash
The open-source ElevenLabs alternative for local voice cloning, design, create, dubbing and dictation Desktop App
TalAter
💬 Speech recognition for your site
ahmetoner
OpenAI Whisper ASR Webservice API
wiltodelta
AI watermark remover. CLI and Python library to strip visible and invisible AI watermarks (Gemini / Nano Banana sparkle, SynthID) and provenance metadata (C2PA, EXIF, IP…
StevenBlack
🔒 Consolidating and extending hosts files from several well-curated sources. Optionally pick extensions for porn, social media, and other categories.
enescingoz
280+ free n8n automation templates — ready-to-use workflows for Gmail, Telegram, Slack, Discord, WhatsApp, Google Drive, Notion, OpenAI, and more. AI agents, RAG chatbot…
iawia002
👾 Fast and simple video download library and CLI tool written in Go
slymnoyann
Hey is a decentralized and permissionless social media app built with Lens Protocol 🌿
wppconnect-team
WPPConnect is an open source project developed by the JavaScript community with the aim of exporting functions from WhatsApp Web to the node, which can be used to suppor…
wuyoscar
GPT Image 2 prompt gallery, image prompt library, agentic skill, and CLI for OpenAI image generation/editing
libvips
A fast image processing library with low memory needs.
lancedb
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
pymupdf
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
Pluviobyte
Build a reproducible local rough-cut workflow for talking-head or narrated screen-recording videos with transcript review gates.