CICD For Machine Learning
kingabzpro
A beginner's project on automating the training, evaluation, versioning, and deployment of models using GitHub Actions.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 13 · 16 shown · 222 public entries
Results: 222
kingabzpro
A beginner's project on automating the training, evaluation, versioning, and deployment of models using GitHub Actions.
redis-developer
Using LlamaIndex, Redis, and OpenAI to chat with PDF documents. Supplementary material for blog post on Microsoft Developer Blog
agentmarkup
Install, configure, audit, and fix AgentMarkup machine-readable website metadata in JavaScript web repos. Use when adding @agentmarkup/vite, @agentmarkup/astro, @agentma…
kolenaIO
Rank LLMs, RAG systems, and prompts using automated head-to-head evaluation
yecchen
Code and Data for "MIRAI: Evaluating LLM Agents for Event Forecasting"
rondoflow
Behavioral guidelines to reduce common LLM coding mistakes. Use when writing, reviewing, or refactoring code to avoid overcomplication, make surgical changes, surface as…
atukunare
Turn chat-pasted text/links into a verified, translated, searchable wiki knowledge base. When a user pastes a link or text (in ANY channel — Discord, Slack, CLI, etc.),…
SigNoz
Initialize or repair SigNoz MCP server configuration for Claude Code, Codex, Cursor, VS Code/GitHub Copilot, Claude Desktop, Gemini CLI, Devin CLI, Grok Build, Windsurf,…
GulajavaMinistudio
Behavioral guidelines to reduce common LLM coding mistakes. Use when writing, reviewing, or refactoring code to avoid overcomplication, make surgical changes, surface as…
runwayml
Foundation for building, modifying, debugging, or verifying Runway Dev Platform integrations in an application: connect Dev MCP, use llms.txt to find current resources,…
ysyecust
Multi-agent orchestration using dmux (tmux pane manager for AI agents). Patterns for parallel agent workflows across Claude Code, Codex, OpenCode, and other harnesses. U…
Analyze code changes and update KNOWLEDGE_BASE.md with architectural and feature changes.
ContextJet-ai
Use this when adding evaluation to an LLM/agent app - measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger o…
ContextJet-ai
Use this to build a good evaluation dataset for an LLM app, the part everyone underestimates. Trigger on "make an eval set", "what should I test my LLM on", "I don't hav…
llopresto87
Maintain a project-local, version-pinned wiki of every external dependency. Use whenever a new library is being added, an existing one is being upgraded, an idiom for us…
llopresto87
Prove that a knowledge base actually works before trusting it — after building or adopting docs, a knowledge graph, or a wiki, verify it can orient a fresh agent and res…