Vllm Omni
vllm-project
A framework for efficient model inference with omni-modality models
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
161–176 / 310
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 310
vllm-project
A framework for efficient model inference with omni-modality models
remsky
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model w/multiplatform CPU, AMD, NVIDIA GPU PyTorch support, handling, and auto-stitching
denizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
Picovoice
On-device wake word detection powered by deep learning
HVision-NKU
Official implementation of ImageCritic (CVPR 2026)
bytedance-fanqie-ai
[ICLR 2026]🔥🔥🔥MOSAIC: Multi-Subject Personalized Generation via Correspondence-Aware Alignment and Disentanglement
francozanardi
A Python library for efficient image generation using CSS Flexbox
aHapBean
[NeurIPS 2025] VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models
k2-fsa
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows,…
IAHispano
A simple, high-quality voice conversion tool focused on ease of use and performance.
World-In-World
Code implementation of the paper "World-in-World: World Models in a Closed-Loop World" (ICLR'26 Oral)
PaddlePaddle
PaddlePaddle GAN library, including lots of interesting applications like First-Order motion transfer, Wav2Lip, picture repair, image editing, photo2cartoon, image style…
abhiTronix
A High-performance cross-platform Video Processing Python framework powerpacked with unique trailblazing features :fire:
myshell-ai
High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.