No visual example yet
Explore the skillFara
microsoft
Fara-7B: An Efficient Agentic Model for Computer Use
OPENAGENTSKILL / DIRECTORY
Temukan skill untuk tugas berikutnya dengan Codex, Claude Code, Cursor, dan lainnya.
97 Skills
Hasil: 97
No visual example yet
Explore the skillmicrosoft
Fara-7B: An Efficient Agentic Model for Computer Use
No visual example yet
Explore the skillNirantK
Curated list of Machine Learning, NLP, Vision, Recommender Systems Project Ideas
No visual example yet
Explore the skillicereed
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
No visual example yet
Explore the skillmagnitudedev
Open-source, vision-first browser agent
No visual example yet
Explore the skillfikrikarim
On-device, real-time multimodal AI. Have natural voice and vision conversations with an AI that runs entirely on your machine. Powered by Gemma 4 E2B and Kokoro.
No visual example yet
Explore the skillNVIDIA-AI-Blueprints
The NVIDIA VSS Blueprint is a suite of reference architectures for building GPU-accelerated vision agents and AI-powered video analytics applications.
No visual example yet
Explore the skillA comprehensive list of Deep Learning / Artificial Intelligence and Machine Learning tutorials - rapidly expanding into areas of AI/Deep Learning / Machine Vision / NLP…
No visual example yet
Explore the skillARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
No visual example yet
Explore the skillsiddsachar
Row-Bot - Personal AI Sovereignty. A local-first AI assistant with integrated tools, a personal knowledge graph, voice, vision, shell, browser automation, scheduled task…
No visual example yet
Explore the skillTurbo1123
Android Automation Tool Based on Vision-Language Models
No visual example yet
Explore the skillemcf
Get clean data from tricky documents, powered by vision-language models ⚡
No visual example yet
Explore the skillsparklabx
Teach your AI to draw correct, beautiful draw.io diagrams — declarative layout engine, ground-truth stencils, structural validator, vision self-check. AWS · Azure · GCP…
No visual example yet
Explore the skilltanelpoder
0x.Tools: X-Ray vision for Linux systems
No visual example yet
Explore the skilltrycua
A curated list of resources about AI agents for Computer Use, including research papers, projects, frameworks, and tools.
No visual example yet
Explore the skillkairyou
Reusable Agent Skills, plus integrations (statusline, provider usage, vision) that install into Codex, Claude Code, and opencode.
No visual example yet
Explore the skillTIGER-AI-Lab
This repo contains the code for "VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks" [ICLR 2025]