技能目录

为 AI Agent 发现可复用技能。

按任务搜索真实的 GitHub 技能,并在使用前查看 Stars、信任、审计、分类和安装路径。

每个推荐都保留与其仓库、审计和安装路径的明确关联。

搜索结果: speech-service

英文目录

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

23K
Stars
87/100
信任
分类: media-automation审计
Cli75

A versatile command-line tool for interacting with Google Workspace APIs, designed for both human users and AI agents.

30K
Stars
75/100
信任
分类: utility审计

Git with a cup of tea! Painless self-hosted all-in-one software development service, including Git hosting, code review, team collaboration, package registry and CI/CD

56K
Stars
86/100
信任
分类: github-automation审计

Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

15K
Stars
87/100
信任
分类: media-automation审计

Open-source live-chat, email support, omni-channel desk. An alternative to Intercom, Zendesk, Salesforce Service Cloud etc. 🔥💬

32K
Stars
78/100
信任
分类: support-automation审计

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

31K
Stars
87/100
信任
分类: media-automation审计

Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.

8.6K
Stars
77/100
信任
分类: media-automation审计

An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

21K
Stars
77/100
信任
分类: media-automation审计

Temporal service

21K
Stars
82/100
信任
分类: automation审计
Iii77

Effortlessly compose, extend, and observe every service in real-time for the first time ever.

18K
Stars
77/100
信任
分类: coding-agents审计

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages

13K
Stars
87/100
信任
分类: media-automation审计

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

13K
Stars
87/100
信任
分类: media-automation审计