Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
$ npx skills add open-compass/VLMEvalKitAlternatives
Compare similar skills by workflow fit, trust score, quality, GitHub adoption, maintenance, and install readiness.
Current skill
Florence-2 is a novel vision foundation model with a unified, prompt-based representation for a variety of computer vision and vision-language tasks.
Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
$ npx skills add open-compass/VLMEvalKitAI watermark remover. CLI and Python library to strip visible and invisible AI watermarks (Gemini / Nano Banana sparkle, SynthID) and provenance metadata (C2PA, EXIF, IPTC) from images.
$ npx skills add wiltodelta/remove-ai-watermarks3D Computer Vision Framework
$ npx skills add alicevision/AliceVisionAI IDE for hardware development, support Arduino, MicroPython, ESP32, STM32, RP2040, Nrf5x...
$ npx skills add ailyProject/aily-blockly๐ค The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
$ npx skills add huggingface/datasets12 Weeks, 24 Lessons, AI for All!
$ npx skills add microsoft/AI-For-BeginnersA collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3, and Qwen3-VL.
$ npx skills add roboflow/notebooksWe write your reusable computer vision tools. ๐
$ npx skills add roboflow/supervisionCross-platform, customizable ML solutions for live and streaming media.
$ npx skills add google-ai-edge/mediapipeUltralytics YOLO ๐
$ npx skills add ultralytics/ultralyticsAdvanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.
$ npx skills add jacobgil/pytorch-grad-camLow-code framework for building custom LLMs, neural networks, and other AI models
$ npx skills add ludwig-ai/ludwigCLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
$ npx skills add openai/CLIPFast and Accurate ML in 3 Lines of Code
$ npx skills add autogluon/autogluon๐ Geometric Computer Vision Library for Spatial AI
$ npx skills add kornia/korniaโ๏ธ DEPRECATED โ See https://github.com/ageron/handson-ml3 or handson-mlp instead.
$ npx skills add ageron/handson-mlHow to choose
Use an alternative when it has a clearer install path, higher trust score, fresher maintenance, or better platform fit for your current agent stack. Keep Florence 2 Vision Language Model if it already passes your workflow test and repository review.
Next step
Open the compare page, test the install commands in a sandbox, and check each repository before using a skill in production.
Alternatives
Compare similar skills by workflow fit, trust score, quality, GitHub adoption, maintenance, and install readiness.
Current skill
Florence-2 is a novel vision foundation model with a unified, prompt-based representation for a variety of computer vision and vision-language tasks.
Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
$ npx skills add open-compass/VLMEvalKitAI watermark remover. CLI and Python library to strip visible and invisible AI watermarks (Gemini / Nano Banana sparkle, SynthID) and provenance metadata (C2PA, EXIF, IPTC) from images.
$ npx skills add wiltodelta/remove-ai-watermarks3D Computer Vision Framework
$ npx skills add alicevision/AliceVisionAI IDE for hardware development, support Arduino, MicroPython, ESP32, STM32, RP2040, Nrf5x...
$ npx skills add ailyProject/aily-blockly๐ค The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
$ npx skills add huggingface/datasets12 Weeks, 24 Lessons, AI for All!
$ npx skills add microsoft/AI-For-BeginnersA collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3, and Qwen3-VL.
$ npx skills add roboflow/notebooksWe write your reusable computer vision tools. ๐
$ npx skills add roboflow/supervisionCross-platform, customizable ML solutions for live and streaming media.
$ npx skills add google-ai-edge/mediapipeUltralytics YOLO ๐
$ npx skills add ultralytics/ultralyticsAdvanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.
$ npx skills add jacobgil/pytorch-grad-camLow-code framework for building custom LLMs, neural networks, and other AI models
$ npx skills add ludwig-ai/ludwigCLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
$ npx skills add openai/CLIPFast and Accurate ML in 3 Lines of Code
$ npx skills add autogluon/autogluon๐ Geometric Computer Vision Library for Spatial AI
$ npx skills add kornia/korniaโ๏ธ DEPRECATED โ See https://github.com/ageron/handson-ml3 or handson-mlp instead.
$ npx skills add ageron/handson-mlHow to choose
Use an alternative when it has a clearer install path, higher trust score, fresher maintenance, or better platform fit for your current agent stack. Keep Florence 2 Vision Language Model if it already passes your workflow test and repository review.
Next step
Open the compare page, test the install commands in a sandbox, and check each repository before using a skill in production.