No visual example yet
Explore the skillAwesome Grounding
TheShadow29
awesome grounding: A curated list of research papers in visual grounding
OPENAGENTSKILL / DIRECTORY
다음 작업에 맞는 스킬을 찾아보세요. Codex, Claude Code, Cursor 등을 지원합니다.
25 Skills
검색 결과: 25
No visual example yet
Explore the skillTheShadow29
awesome grounding: A curated list of research papers in visual grounding
No visual example yet
Explore the skillmbzuai-oryx
[CVPR 2024 🔥] Grounding Large Multimodal Model (GLaMM), the first-of-its-kind model capable of generating natural language responses that are seamlessly integrated with…
No visual example yet
Explore the skillAnionex
给纯文本 LLM agent 装上眼睛:图片问答、OCR、截图分析、视觉定位等一套视觉工具箱 + skill,并可无缝接入 Codex、Claude Code、OpenCode、Pi | Give text-only LLM agents vision: image Q&A, OCR, screenshot understanding,…
No visual example yet
Explore the skillAnionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
No visual example yet
Explore the skilltsingyuai
使用 Bing Webmaster 页面与查询数据、索引状态、产品结果和 Bing AI Performance 证据复盘 SEO 页面,诊断收录、排名、点击率、意图、内容、转化和 AI 引用问题。需要比较周期表现、分析 citations、cited pages、grounding queries、query fan-out 或决定下…
No visual example yet
Explore the skillNVIDIA-AI-Blueprints
Generates VSS video summary reports with LVS HITL and optional Enterprise RAG document grounding. Trigger when the user asks for a frag/RAG-assisted video report, knowle…
No visual example yet
Explore the skilltjboudreaux
Use when a specific claim may lack grounding. Check evidence boundary, size wrongness cost, then answer, fetch, or abstain — never confabulate.
No visual example yet
Explore the skillImplement a production GTAO path in Three.js. Use for half-resolution horizon sampling, reversed-depth reconstruction, bent-normal encoding, full-resolution bilateral re…
No visual example yet
Explore the skillVarnan-Tech
Use when the user asks to generate a blog cover image, thumbnail, or article header. Automatically uses modern typography, brand logos, and Google Search grounding to cr…
No visual example yet
Explore the skillkohjingyu
🧀 Code and models for the ICML 2023 paper "Grounding Language Models to Images for Multimodal Inputs and Outputs".
No visual example yet
Explore the skillsuitedaces
Generate and edit images using the Gemini API. Text-to-image, image editing, multi-turn iteration, 4K resolution, search grounding.
No visual example yet
Explore the skillCamusGIT
Quant-focused research ideation pipeline: scope selection (3 stages) → anchor-first literature grounding → single-core idea generation → iterative refinement → ELO tourn…
No visual example yet
Explore the skillagentii-ai
Dated FDA catalyst calendar (PDUFA target dates, AdCom meetings, device decisions, trial readouts, earnings) with AdCom-style scrutiny outcome framing and historical-cas…
No visual example yet
Explore the skillagentii-ai
Clinical-trial readout analysis: pull the trial, evaluate the readout with AdCom-style scrutiny (endpoints, statistics, subgroups, missing data, safety), and size the st…
No visual example yet
Explore the skillRaidriar7170
Vision-grounding plugin for browser-use agents with SoM, Florence-2, Vision–DOM alignment, adaptive visual context, and objective evaluation.
No visual example yet
Explore the skillK-Dense-AI
Applies the empiricist, skeptical, and practical reasoning of David Hume, 18th-century Scottish philosopher. Use this skill whenever you encounter questions about episte…