LlamaGen
FoundationVision
Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 16
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 16
FoundationVision
Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation
leejet
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
PABannier
Suno AI's Bark model in C/C++ for fast text-to-speech generation
unslothai
Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
k2-fsa
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Su…
supertone-inc
Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.
flashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
k2-fsa
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows,…
k2-fsa
Speech-to-text server framework with next-gen Kaldi
fluxions-ai
Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime o…
zhenye234
LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesis
PurpleDoubleD
Local AI desktop app — chat, agent mode, image gen, video gen. Supports Ollama, Gemma 4, Llama, Qwen, OpenAI, Anthropic. Single .exe, no Docker.
intelligentnode
Access the latest AI models like ChatGPT, LLaMA, Deepseek, Diffusion, Hugging face, and beyond through a unified prompt layer and performance evaluation