Directorio de skills

Descubre skills reutilizables para AI agents.

Busca skills reales de GitHub por tarea y revisa stars, confianza, auditoría, categoría y ruta de instalación antes de utilizarlos.

Cada recomendación conserva un vínculo claro con su repositorio, auditoría y ruta de instalación.

Resultados de búsqueda: frontier

Directorio en inglés

AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs

50K
Stars
88/100
Confianza
Categoría: agent-skillsAuditoría

Financial portfolio optimisation in python, including classical efficient frontier, Black-Litterman, Hierarchical Risk Parity

5.8K
Stars
81/100
Confianza
Categoría: financeAuditoría

Vowpal Wabbit is a machine learning system which pushes the frontier of machine learning with techniques such as online, hashing, allreduce, reductions, learning2search, active, and interactive learning.

8.7K
Stars
76/100
Confianza
Categoría: ml-automationAuditoría

🚀 MassGen is an open-source multi-agent scaling system that runs in your terminal, autonomously orchestrating frontier models and agents to collaborate, reason, and produce high-quality results. | Join us on Discord: discord.massgen.ai

1.1K
Stars
73/100
Confianza
Categoría: agent-frameworksAuditoría

Autonomously improve a real artifact (code, training recipe, agent harness, data pipeline, prompt) against an objective and an evaluator, using Hypothesis Tree Refinement (HTR) from the Arbor paper. Use this whenever someone wants to iteratively optimize something over many experiments without overfitting — e.g. "get my model's eval score up", "improve this agent/harness", "tune this pipeline", "beat the baseline on this benchmark", "run a search over approaches and keep the best", "do an MLE-bench / Kaggle-style optimization", or any long-horizon "make this artifact better and don't just memorize the dev set" task. Trigger it even when the user doesn't say "Arbor" or "hypothesis tree" but describes repeated experiment-and-evaluate loops, branching exploration of competing ideas, or worries about a dev/test gap. Runs Claude itself as the coordinator with subagent executors in isolated git worktrees; for the standalone `arbor` CLI tool see references/arbor-upstream.md.

34K
Stars
77/100
Confianza
Categoría: researchAuditoría

Measuring frontier coding agents on original, long-horizon engineering tasks

944
Stars
67/100
Confianza
Categoría: coding-agentsAuditoría

MathCode: A Frontier Mathematical Coding Agent

575
Stars
64/100
Confianza
Categoría: coding-agentsAuditoría

Catch your AI's mistakes and blind spots before your customers or regulators do. iFixAi runs 45 inspections, 32 graded core plus 13 extended for frontier risks like sabotage, sandbagging, and oversight evasion. It returns a letter grade in under 5 minutes. Industry and model agnostic.

479
Stars
67/100
Confianza
Categoría: financeAuditoría

The only real free CLI agent. Harvests your Gemini (guest, no login) · Claude.ai · Claude Code · Kimi · Qwen · DeepSeek browser session and turns it into a tool-calling agent — reads & edits files, runs Bash, greps your repo, browses the web, ships commits, all from your terminal. Frontier IA driving real work at $0

379
Stars
64/100
Confianza
Categoría: developmentAuditoría

"QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks"

168
Stars
68/100
Confianza
Categoría: researchAuditoría

Curated marketplace of AI skills, agents, and rules for cloud, zero-trust, and compliance-aware engineering - works with Claude Code, Codex, Cursor, Copilot, and more.

20
Stars
66/100
Confianza
Categoría: utilityAuditoría

Claude Code skill plugin for cleaning up and bringing Lean 4 code to mathlib standards.

27
Stars
66/100
Confianza
Categoría: coding-agentsAuditoría