Mlx Tune
ARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning β natively on MLX. Unsloth-compatible API.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
17β32 / 71
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 71
ARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning β natively on MLX. Unsloth-compatible API.
SandAI-org
MAGI-1: Autoregressive Video Generation at Scale
nvidia-cosmos
Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the future state of the world in theβ¦
shivammehta25
[ICASSP 2024] π΅ Matcha-TTS: A fast TTS architecture with conditional flow matching
Tencent
High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
bytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
FoundationVision
[CVPR 2025 Oral]Infinity β : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
jy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
lifeiteng
PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo https://lifeiteng.github.io/valle/index.html
PKU-YuanGroup
[TPAMI 2025π₯] MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
FireRedTeam
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity geneβ¦
Phantom-video
Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment
FoundationVision
Autoregressive Model Beats Diffusion: π¦ Llama for Scalable Image Generation
Tencent-Hunyuan
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
wladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
menyifang
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"