SageAttention
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
Media AutomationReview before use
3.4KGitHub
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–3 / 3
Results: 3
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
thu-ml
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
abhiTronix
A High-performance cross-platform Video Processing Python framework powerpacked with unique trailblazing features :fire:
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.