SageAttention
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–7 / 7
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 7
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
thu-ml
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
supertone-inc
Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.
shivammehta25
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
tnfe
A fast video processing library based on node.js (一个基于node.js的高速视频制作库)
DigitalPhonetics
Controllable and fast Text-to-Speech for over 7000 languages!