OnnxStream
vitoplantamura
Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but also Mistral 7B on desktops and…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
17–25 / 25
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 25
vitoplantamura
Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but also Mistral 7B on desktops and…
alibaba
MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI.
meta-llama
Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to e…
NVIDIA
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference appl…
obss
Framework agnostic sliced/tiled inference + interactive ui + error analysis plots
NVIDIA
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, t…
jolibrain
Deep Learning Server and CLI for Torch and TensorRT
4paradigm
OpenMLDB is an open-source machine learning database that provides a feature platform computing consistent features for training and inference.
probcomp
A general-purpose probabilistic programming system with programmable inference