No visual example yet
Explore the skillDiT Extrapolation
thu-ml
Official implementation for "RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers" (ICML 2025) , UltraViCo (ICLR 2026) and UltraImage
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
8 Skills
Results: 8
No visual example yet
Explore the skillthu-ml
Official implementation for "RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers" (ICML 2025) , UltraViCo (ICLR 2026) and UltraImage
No visual example yet
Explore the skillbytedance
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
No visual example yet
Explore the skilljy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
No visual example yet
Explore the skillNVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
No visual example yet
Explore the skillHBAI-Ltd
Toonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI 编剧、智能分镜、角色与视频生成,跨平台桌面端轻量部署,助力创作者低成本批量产出视觉内容。Toonflow is an open-source AI tool that turns stories and scripts into animat…
No visual example yet
Explore the skillBlaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
No visual example yet
Explore the skillvllm-project
A framework for efficient model inference with omni-modality models
No visual example yet
Explore the skilllucidrains
Implementation of Recurrent Interface Network (RIN), for highly efficient generation of images and video without cascading networks, in Pytorch