No visual example yet
Explore the skillUnisondb
ankur-anand
A streaming multimodal database for Edge AI, and Edge Computing.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
65–80 / 84
Results: 84
No visual example yet
Explore the skillankur-anand
A streaming multimodal database for Edge AI, and Edge Computing.
No visual example yet
Explore the skilljaketae
Multimodal AI Story Teller, built with Stable Diffusion, GPT, and neural text-to-speech
No visual example yet
Explore the skillbowang-lab
BioReason: Incentivizing Multimodal Biological Reasoning within a DNA-LLM Model | NeurIPS '25
No visual example yet
Explore the skillkohjingyu
🧀 Code and models for the ICML 2023 paper "Grounding Language Models to Images for Multimodal Inputs and Outputs".
No visual example yet
Explore the skillkohjingyu
🐟 Code and models for the NeurIPS 2023 paper "Generating Images with Multimodal Language Models".
No visual example yet
Explore the skillopengeos
A multimodal AI agent for geospatial data analysis and interactive visualization
No visual example yet
Explore the skillJIA-Lab-research
This project is the official implementation of 'LLMGA: Multimodal Large Language Model based Generation Assistant', ECCV2024 Oral
No visual example yet
Explore the skillWisconsinAIVision
[CVPR2024] ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts
No visual example yet
Explore the skillidwts
🦦 Crayotter: A Multimodal AI-Agent for Video-Editing, Video-Composing, and Video Production. Powered by Multimodal LLMs for autonomous Text-to-Video agentic framework.…
No visual example yet
Explore the skillwanshuiyin
Agentic, long-horizon visual generation: a fuzzy story → a cross-model-audited image-based movie. Brings ARIS's research-wiki + multi-agent debate to multimodal generati…
No visual example yet
Explore the skillveniceai
Call POST /chat/completions on Venice. Covers the OpenAI-compatible request shape, Venice-only venice_parameters (web search, E2EE, characters, thinking control, X searc…
No visual example yet
Explore the skillfirebase
Official skill for integrating Firebase AI Logic (Gemini API) into web applications. Covers setup, multimodal inference, structured output, and security.
No visual example yet
Explore the skillJason904
Design banners for social media, ads, website heroes, creative assets, and print. Multiple art direction options with AI-generated visuals. Actions: design, create, gene…
No visual example yet
Explore the skillmbzuai-oryx
[CVPR 2025 🔥]A Large Multimodal Model for Pixel-Level Visual Grounding in Videos
No visual example yet
Explore the skillDjangoPeng
This repository is a hub for AI Agent projects, including GitHub Sentinel, LanguageMentor, and ChatPPT, designed to enhance enterprise workflows, language learning, and…
No visual example yet
Explore the skillsunshine-lang
从多模态素材到选题、核验、脚本、配图和小红书发布文案的 Codex Skill