No visual example yet
Explore the skillDeep Learning Drizzle
kmario23
Drench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!
OPENAGENTSKILL / DIRECTORY
Temukan skill untuk tugas berikutnya dengan Codex, Claude Code, Cursor, dan lainnya.
97 Skills
Hasil: 97
No visual example yet
Explore the skillkmario23
Drench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!
No visual example yet
Explore the skillkatanaml
Structured data extraction and instruction calling with ML, LLM and Vision LLM
No visual example yet
Explore the skilldmMaze
深度学习辅助漫画翻译工具, 支持一键机翻和简单的图像/文本编辑 | Yet another computer-aided comic/manga translation tool powered by deeplearning
No visual example yet
Explore the skilldusty-nv
Hello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.
No visual example yet
Explore the skillroboflow
computer vision and sports
No visual example yet
Explore the skillGetStream
Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses Stream's edge network for ultra-low latency.
No visual example yet
Explore the skillcambrian-mllm
Cambrian-1 is a family of multimodal LLMs with a vision-centric design.
No visual example yet
Explore the skillAnionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
No visual example yet
Explore the skillshowlab
[CVPR 2025] Open-source, End-to-end, Vision-Language-Action model for GUI Agent & Computer Use.
No visual example yet
Explore the skillkornia
🦀 Low-level 3D Computer Vision library in Rust
No visual example yet
Explore the skillOthersideAI
A framework to enable a multimodal model to operate a computer.
No visual example yet
Explore the skillweb-infra-dev
AI-powered, vision-driven UI automation for every platform.
No visual example yet
Explore the skillAnionex
给纯文本 LLM agent 装上眼睛:图片问答、OCR、截图分析、视觉定位等一套视觉工具箱 + skill,并可无缝接入 Codex、Claude Code、OpenCode、Pi | Give text-only LLM agents vision: image Q&A, OCR, screenshot understanding,…
No visual example yet
Explore the skillAmberSahdev
Control Any Computer Using LLMs.
No visual example yet
Explore the skillgo-vgo
RobotGo, Go Native cross-platform RPA, GUI automation, Auto test and Computer use @vcaesar
No visual example yet
Explore the skillcoasty-ai
State of the Art 82% OSWorld Computer Using Agent, production-ready. Remote and Local! Setup using one API Key