Parlor
fikrikarim
On-device, real-time multimodal AI. Have natural voice and vision conversations with an AI that runs entirely on your machine. Powered by Gemma 4 E2B and Kokoro.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 19
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 19
fikrikarim
On-device, real-time multimodal AI. Have natural voice and vision conversations with an AI that runs entirely on your machine. Powered by Gemma 4 E2B and Kokoro.
ARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
openvinotoolkit
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
junyanz
Image-to-Image Translation in PyTorch
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
NVIDIA
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
junyanz
Software that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.
phillipi
Image-to-image translation with conditional adversarial nets
fudan-generative-vision
[ECCV 2024] Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
open-mmlab
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion mod…
MetalPetal
A GPU accelerated image and video processing framework built on Metal.
thu-ml
Official code of Motus: A Unified Latent Action World Model
cubist38
A high-performance API server that provides OpenAI-compatible endpoints for MLX models. Developed using Python and powered by the FastAPI framework, it provides an effic…
soniqo
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
TheStageAI
Optimized Whisper models for streaming and on-device use
Eyeline-Labs
Official code, models, and data for Vista4D: Video Reshooting with 4D Point Clouds (CVPR 2026 Highlight)