SageAttention
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–8 / 8
Results: 8
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
jy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
thu-ml
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
index-tts
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
NVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
vllm-project
A framework for efficient model inference with omni-modality models
jik876
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.