No visual example yet
Explore the skillVideo Analyzer
byjlw
Analyze videos using LLMs, Computer Vision and Automatic Speech Recognition
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
177–192 / 466
Results: 466
No visual example yet
Explore the skillbyjlw
Analyze videos using LLMs, Computer Vision and Automatic Speech Recognition
No visual example yet
Explore the skilledwko
Interface for OuteTTS models.
No visual example yet
Explore the skillspotify
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection
No visual example yet
Explore the skillR3gm
Synchronized Translation for Videos. Video dubbing
No visual example yet
Explore the skillLucklySpace
A cross-platform instant messaging client application built with Tauri and Vue 3, featuring one-to-one chat, group chat, file transfer, audio/video calling, screen recor…
No visual example yet
Explore the skillantgroup
[CVPR 2025] EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation
No visual example yet
Explore the skilljianchang512
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
No visual example yet
Explore the skillgradio-app
The python library for real-time communication
No visual example yet
Explore the skillyeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…
No visual example yet
Explore the skillmozilla
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
No visual example yet
Explore the skillardha27
AI Vtuber for Streaming on Youtube/Twitch
No visual example yet
Explore the skillgitcoffee-os
PostBot 内容同步助手 一款开源的多平台内容同步分发生产力工具。 支持将文章、笔记、动态、图片、视频、音频等内容,一键同步发布至主流媒体平台。覆盖微信/微博/今日头条/小红书/知乎/百家号/企鹅号/视频号/抖音/快手/哔哩哔哩(B站)等国内主流媒体平台,可轻松扩展兼容 X(Twitter)、Facebook、Instagram、T…
No visual example yet
Explore the skillAzure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
No visual example yet
Explore the skillAratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
No visual example yet
Explore the skillPurfview
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
No visual example yet
Explore the skilllinto-ai
Multilingual Automatic Speech Recognition with word-level timestamps and confidence
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Stock analysis, market research, quant backtesting, financial data, and investment research skills.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.
Choose React motion graphics, generated footage or editing; compare dependencies and inspect actual outputs.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.