PySceneDetect
Breakthrough
:movie_camera: Python and OpenCV-based scene cut/transition detection program & library.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 40
Results: 40
Breakthrough
:movie_camera: Python and OpenCV-based scene cut/transition detection program & library.
tstanislawek
A curated list of resources for Document Understanding (DU) topic
hustvl
[CVPR 2024] 4D Gaussian Splatting for Real-Time Dynamic Scene Rendering
bytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
godot-gdunit-labs
Embedded unit testing framework for Godot 4 supporting GDScript and C#. Features test-driven development, embedded test inspector, extensive assertions, mocking, scene t…
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
google-labs-code
A format specification for describing a visual identity to coding agents. DESIGN.md gives agents a persistent, structured understanding of a design system.
Countly
Countly is a privacy-first, AI-powered analytics and engagement platform for understanding and optimizing customer journeys across digital applications, from desktop and…
Mai-with-u
MaiSaka, an LLM-based intelligent agent, is a digital lifeform devoted to understanding you and interacting in the style of a real human. She does not pursue perfection,…
oceanbase
MiniOB is a compact database that assists developers in understanding the fundamental workings of a database.
pguso
Demystify AI agents by building them yourself. Local LLMs, no black boxes, real understanding of function calling, memory, and ReAct patterns.
SublimeText-Markdown
Powerful Markdown package for Sublime Text with better syntax understanding and good color schemes.
clovaai
Official Implementation of OCR-free Document Understanding Transformer (Donut) and Synthetic Document Generator (SynthDoG), ECCV 2022
Anionex
给纯文本 LLM agent 装上眼睛:图片问答、OCR、截图分析、视觉定位等一套视觉工具箱 + skill,并可无缝接入 Codex、Claude Code、OpenCode、Pi | Give text-only LLM agents vision: image Q&A, OCR, screenshot understanding,…
open-mmlab
OpenMMLab Text Detection, Recognition and Understanding Toolbox
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.