InfinityStar
FoundationVision
[NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
113–128 / 430
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 430
Live data is unavailable. A saved snapshot may be shown; check the source before use.
FoundationVision
[NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation
intel
Libva is an implementation for VA-API (Video Acceleration API)
X-GenGroup
A unified framework for easy reinforcement learning in Flow-Matching models
Eyeline-Labs
Official code, models, and data for Vista4D: Video Reshooting with 4D Point Clouds (CVPR 2026 Highlight)
sipeter
A lightweight, offline Android Text-to-Speech (TTS) engine enabling seamless system-wide voice cloning and high-fidelity text reading. / 运行在安卓本地的轻量级文字转语音 (TTS) 引擎,支持离线发音…
salute-developers
Foundational Model for Speech Recognition Tasks
tmchow
illo skill — an AI agent skill that turns ideas and articles into original print-style editorial illustrations, starring a recurring mascot. 30+ characters packs, with a…
mirabarukaso
Character Select Stand Alone App with AI prompt and ComfyUI/WebUI API support for wai-il model
SaladTechnologies
A simple API server to make ComfyUI easy to scale horizontally. Get outputs directly in the response, or asynchronously via a variety of storage providers.
elevenlabs
The official JavaScript (Node) library for the ElevenLabs API.
Saganaki22
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
nv-tlabs
Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models
zarazhangrui
Turn any content into a personalized AI podcast. NotebookLM-style, except you control the script, voices, and hosts. Listen in Apple Podcasts, Spotify, or any podcast ap…
wernerturing
Detect and remove logos from videos, even if they change position several times
dectalk
Modern builds for the 90s/00s DECtalk text-to-speech application.