No visual example yet
Explore the skillVision Agents
GetStream
Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses Stream's edge network for ultra-low latency.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
92 Skills
Results: 92
No visual example yet
Explore the skillGetStream
Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses Stream's edge network for ultra-low latency.
No visual example yet
Explore the skillalicevision
3D Computer Vision Framework
No visual example yet
Explore the skillazavea
An open source library and framework for deep learning on satellite and aerial imagery.
No visual example yet
Explore the skillCharmve
A computer vision closed-loop learning platform where code can be run interactively online. 学习闭环《计算机视觉实战演练:算法与应用》中文电子书、源码、读者交流社区(持续更新中 ...) 📘 在线电子书 https://charmve.gith…
No visual example yet
Explore the skillcleanlab
Automatically find issues in image datasets and practice data-centric computer vision.
No visual example yet
Explore the skillDirtyHarryLYL
Recent Transformer-based CV and related works.
No visual example yet
Explore the skillunrealcv
A list of synthetic dataset and tools for computer vision
No visual example yet
Explore the skillpytorch
Datasets, Transforms and Models specific to Computer Vision
No visual example yet
Explore the skillhuggingface
This repo is the homebase of a community driven course on Computer Vision with Neural Networks. Feel free to join us on the Hugging Face discord: hf.co/join/discord
No visual example yet
Explore the skillroboflow
We write your reusable computer vision tools. 💜
No visual example yet
Explore the skillmicrosoft
Best Practices, code samples, and documentation for Computer Vision.
No visual example yet
Explore the skillNo visual example yet
Explore the skillAnionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
No visual example yet
Explore the skillautowarefoundation
Free self-driving car stack - fully open-source ADAS and autonomous driving system
No visual example yet
Explore the skillxiincs
为 Claude Code 赋能多模态视觉能力,支持豆包、通义千问、GPT-4o 等模型,用于截图 / UI / 图表分析;适配 DeepSeek 等无视觉底座,搭配 browser-harness 可做前端布局自动化检查。
No visual example yet
Explore the skillAnionex
给纯文本 LLM agent 装上眼睛:图片问答、OCR、截图分析、视觉定位等一套视觉工具箱 + skill,并可无缝接入 Codex、Claude Code、OpenCode、Pi | Give text-only LLM agents vision: image Q&A, OCR, screenshot understanding,…