Speech To Text Benchmark
Picovoice
speech to text benchmark framework
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Browse all public registry entries, one page at a time. Counts include resources; MCP-only resources are omitted from displayed skills. Listing is not a safety or compatibility guarantee.
Page 25 · 16 shown · 824 public entries
Results: 824
Picovoice
speech to text benchmark framework
sipeter
A lightweight, offline Android Text-to-Speech (TTS) engine enabling seamless system-wide voice cloning and high-fidelity text reading. / 运行在安卓本地的轻量级文字转语音 (TTS) 引擎,支持离线发音…
vilassn
Offline Speech Recognition with OpenAI Whisper and TensorFlow Lite for Android
Chen-Yang-Liu
[IEEE GRSM 2025 🔥] "Text2Earth: Unlocking Text-driven Remote Sensing Image Generation with a Global-Scale Dataset and a Foundation Model"
sveinbjornt
Command line interface for the built-in speech recognition and transcription capabilities in macOS.
judahpaul16
ChatGPT at home! A better alternative to commercial smart home assistants, built on the Raspberry Pi using LiteLLM and LangGraph.
jeffstric
ZhiJuTong (ZJT) is an AI-powered, open-source platform specifically designed for creating professional short dramas. It automates the entire production pipeline, from sc…
ShandaAI
Generative World Renderer: an AI-native Renderer for Games and Virtual Worlds. 面向游戏与虚拟世界的AI原生渲染引擎
prouast
Desktop implementation of Remote Photoplethysmography – Measuring heart rate using facial video.
salute-developers
Foundational Model for Speech Recognition Tasks
HVision-NKU
Official implementation of ImageCritic (CVPR 2026)
ddalcu
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling.
Bomx
AI video production skill for agents: HeyGen avatars, Seedance b-roll, OpenAI images, Remotion, HyperFrames, screen recording, FFmpeg captions and QC.
gudaochangsheng
[CVPR 2026] Official PyTorch implementation of WaDi: Weight Direction-aware Distillation for One-step Image Synthesis
NJU-3DV
[CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations
mirabarukaso
Character Select Stand Alone App with AI prompt and ComfyUI/WebUI API support for wai-il model
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
RedSkill · 流白Livo · 1.0.0
Describe a rain curtain, growing flowering branches or a flock of swallows. This RedSkill package adapts three p5.js templates into sketch.js code with custom colors, density and speed.
Noncommercial use only · Local runtime recording · Use on the source platform