SageAttention
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 70
Results: 70
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
werf
A solution for implementing efficient and consistent software delivery to Kubernetes facilitating best practices.
takara-ai
A full attention mechanism and transformer in pure go.
josephmachado
Code for "Efficient Data Processing in Spark" Course
mryab
Efficient Deep Learning Systems course materials (HSE, YSDA)
lucidrains
Reformer, the efficient Transformer, in Pytorch
modelscope
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
dotnetcore
DotnetSpider, a .NET standard web crawling library. It is lightweight, efficient and fast high-level web crawling & scraping framework
pulldown-cmark
An efficient, reliable parser for CommonMark, a standard dialect of Markdown
SkyworkAI
DeepResearchAgent is a hierarchical multi-agent system designed not only for deep research tasks but also for general-purpose task solving. The framework leverages a top…
NovaSearch-Team
Unify Efficient Fine-tuning of RAG Retrieval, including Embedding, ColBERT, ReRanker.
jy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
ELS-RD
Efficient, scalable and enterprise-grade CPU/GPU inference server for 🤗 Hugging Face transformer models 🚀
RQLuo
MixTeX multimodal LaTeX, ZhEn, and, Table OCR. It performs efficient CPU-based inference in a local offline on Windows.
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.