No visual example yet
Explore the skillCosyVoice
FunAudioLLM
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
107 Skills
Results: 107
No visual example yet
Explore the skillFunAudioLLM
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
No visual example yet
Explore the skillscreenpipe
YC (S26) | AI that knows what you've seen, said, or heard. Records everything you do, say, hear 24/7, local, private, secure
No visual example yet
Explore the skillxorbitsai
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all throug…
No visual example yet
Explore the skillmodelscope
ModelScope: bring the notion of Model-as-a-Service to life.
No visual example yet
Explore the skillbghira
A general fine-tuning kit geared toward image/video/audio diffusion models.
No visual example yet
Explore the skillspotify
No visual example yet
Explore the skillNVIDIA
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference appl…
No visual example yet
Explore the skillEventual-Inc
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
No visual example yet
Explore the skillwenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
No visual example yet
Explore the skillhuggingface
Build local voice agents with open-source models
No visual example yet
Explore the skillyanshengjia
Machine Learning and Agentic AI Resources, Practice and Research
No visual example yet
Explore the skillwhitphx
Real-time video and audio processing on Streamlit
No visual example yet
Explore the skillOpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
No visual example yet
Explore the skillpytorch
Data manipulation and transformation for audio signal processing, powered by PyTorch
No visual example yet
Explore the skillNatively-AI-assistant
Natively — Free open-source AI meeting assistant, interview copilot, and note taker. The best alternative to Cluely, Otter, Granola, Final Round AI, Fireflies, and Inter…
No visual example yet
Explore the skillmicrosoft
A Unified Semi-Supervised Learning Codebase (NeurIPS'22)
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Stock analysis, market research, quant backtesting, financial data, and investment research skills.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.
Choose React motion graphics, generated footage or editing; compare dependencies and inspect actual outputs.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.