No visual example yet
Explore the skillSparrow
katanaml
Structured data extraction and instruction calling with ML, LLM and Vision LLM
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
92 Skills
Results: 92
No visual example yet
Explore the skillkatanaml
Structured data extraction and instruction calling with ML, LLM and Vision LLM
No visual example yet
Explore the skilldusty-nv
Hello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.
No visual example yet
Explore the skillfudan-generative-vision
[ECCV 2024] Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
No visual example yet
Explore the skillroboflow
No visual example yet
Explore the skillriddleling
An iOS OCR Server Using Apple’s Vision Framework
No visual example yet
Explore the skillshowlab
[CVPR 2025] Open-source, End-to-end, Vision-Language-Action model for GUI Agent & Computer Use.
No visual example yet
Explore the skillcambrian-mllm
Cambrian-1 is a family of multimodal LLMs with a vision-centric design.
No visual example yet
Explore the skillsymisc
An Embedded Computer Vision & Machine Learning Library (CPU Optimized & IoT Capable)
No visual example yet
Explore the skill2toinf
[ICLR 2026] The offical Implementation of "Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model"
No visual example yet
Explore the skillkornia
🦀 Low-level 3D Computer Vision library in Rust
No visual example yet
Explore the skillprivatenumber
macOS CLI for OCR and searchable PDFs using Apple's Vision framework
No visual example yet
Explore the skillYanjieZe
A paper list of my history reading. Robotics, Learning, Vision.
No visual example yet
Explore the skillhuggingface
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and traini…
No visual example yet
Explore the skillAgents365-ai
Generate draw.io diagrams from natural language — 6 presets, vision self-check + up to 5-round refinement, codebase-to-diagram, 10,000+ official shapes & 321 AI/LLM bran…
No visual example yet
Explore the skillsou350121
本项目旨在为致力于进入VLA(Vision-Language-Action)领域的算法工程师提供一份全中文、实战导向的学习/面试手册。 不同于通用的 CV/NLP 面试指南,本项目聚焦于 Robotics 特有的挑战
No visual example yet
Explore the skillweb-infra-dev
AI-powered, vision-driven UI automation for every platform.