暂未收录效果图
查看技能说明Face Recognition
ageitgey
The world's simplest facial recognition api for Python and the command line
OPENAGENTSKILL / DIRECTORY
为下一项任务找到合适的技能。探索适用于 Codex、Claude Code、Cursor 等 Agent 的工具。
38 Skills
搜索结果: 38
暂未收录效果图
查看技能说明ageitgey
The world's simplest facial recognition api for Python and the command line
暂未收录效果图
查看技能说明Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
暂未收录效果图
查看技能说明nl8590687
暂未收录效果图
查看技能说明m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
暂未收录效果图
查看技能说明FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
暂未收录效果图
查看技能说明serengil
A Lightweight Face Recognition and Facial Attribute Analysis (Age, Gender, Emotion and Race) Library for Python
暂未收录效果图
查看技能说明alphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
暂未收录效果图
查看技能说明TalAter
暂未收录效果图
查看技能说明mindee
docTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks powered by Deep Learning.
暂未收录效果图
查看技能说明wenet-e2e
暂未收录效果图
查看技能说明MahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
暂未收录效果图
查看技能说明flashlight
暂未收录效果图
查看技能说明jianchang512
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
暂未收录效果图
查看技能说明FireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
暂未收录效果图
查看技能说明FunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
暂未收录效果图
查看技能说明royshil