No visual example yet
Explore the skillffmpeg
calesthio
Video and audio processing with FFmpeg. Use for format conversion, resizing, compression, audio extraction, and preparing assets for Remotion. Triggers include convertin…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
30 Skills
Results: 30
No visual example yet
Explore the skillcalesthio
Video and audio processing with FFmpeg. Use for format conversion, resizing, compression, audio extraction, and preparing assets for Remotion. Triggers include convertin…
No visual example yet
Explore the skillcalesthio
DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen…
No visual example yet
Explore the skillcalesthio
Generate AI voiceovers, sound effects, and music using ElevenLabs APIs. Use when creating audio content for videos, podcasts, or games. Triggers include generating voice…
No visual example yet
Explore the skillgoogle-ai-edge
Cross-platform, customizable ML solutions for live and streaming media.
No visual example yet
Explore the skillhuggingface
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and traini…
No visual example yet
Explore the skillalyssaxuu
The free and privacy-friendly screen recorder with no limits 🎥
No visual example yet
Explore the skillFunAudioLLM
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
No visual example yet
Explore the skillscreenpipe
YC (S26) | AI that knows what you've seen, said, or heard. Records everything you do, say, hear 24/7, local, private, secure
No visual example yet
Explore the skillxorbitsai
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all throug…
No visual example yet
Explore the skillmodelscope
ModelScope: bring the notion of Model-as-a-Service to life.
No visual example yet
Explore the skillbghira
A general fine-tuning kit geared toward image/video/audio diffusion models.
No visual example yet
Explore the skillspotify
No visual example yet
Explore the skillNVIDIA
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference appl…
No visual example yet
Explore the skillEventual-Inc
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
No visual example yet
Explore the skillhuggingface
Build local voice agents with open-source models
No visual example yet
Explore the skillyanshengjia
Machine Learning and Agentic AI Resources, Practice and Research
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Stock analysis, market research, quant backtesting, financial data, and investment research skills.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.
Choose React motion graphics, generated footage or editing; compare dependencies and inspect actual outputs.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.