Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
$ npx skills add FunAudioLLM/CosyVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
OpenAI Agents + CLI · 4 targets