Echo Memory
Echo-Team-Joy-Future-Academy-JD
A Simple Baseline for Video World Models with Memory
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
193–208 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
Echo-Team-Joy-Future-Academy-JD
A Simple Baseline for Video World Models with Memory
taco-group
SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
OpenDriveLab
[CVPR 2024 Highlight] GenAD: Generalized Predictive Model for Autonomous Driving
NJU-3DV
[CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations
helloianneo
Xiaohei 2.0 Codex Skill for Chinese real-object article illustrations and long-scroll story images
githubharald
Connectionist Temporal Classification (CTC) decoder with dictionary and language model.
Aratako
Multilingual TTS model with voice cloning and duration control, based on T5Gemma encoder-decoder LLM
lucidrains
Implementation of Recurrent Interface Network (RIN), for highly efficient generation of images and video without cascading networks, in Pytorch
Visko-Platform
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
light-and-ray
MCWW: Additional non-node based UI for ComfyUI focused on inference. Stable UI states; presets; and advanced queue. Based on Gradio
FoundationVision
[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation…
unslothai
Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
huggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
OpenBMB
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning