OPENAGENTSKILL / DIRECTORY

AI Agent Skills

Finde den passenden Skill für deine nächste Aufgabe mit Codex, Claude Code, Cursor und mehr.

Ergebnisse · “swe-bench”

14 Skills

Ergebnisse: 14

Entwicklung

No visual example yet

Explore the skill

The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench veri…

Preis unbestätigtEntwicklungVor Nutzung prüfen
5287GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

SE Agent

JARVIS-Xs

SE-Agent is a self-evolution framework for LLM Code agents. It enables trajectory-level evolution to exchange information across reasoning paths via Revision, Recombinat…

Preis unbestätigtEntwicklungClaude CodeVor Nutzung prüfen
280GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

Chronos

Kodezi

Kodezi Chronos is a debugging-first language model that achieves state-of-the-art results on SWE-bench Lite (80.33%) and 67% real-world fix accuracy, over six times bett…

Preis unbestätigtEntwicklungOpenAI AgentsVor Nutzung prüfen
4917GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

An Alibaba open-source multi-language benchmark for evaluating LLMs in repository-level automatic code review, featuring an AI-assisted and expert-verified dataset.

Preis unbestätigtEntwicklungVor Nutzung prüfen
209GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

SWE Agent

SWE-agent

SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding chal…

Preis unbestätigtEntwicklungVor Nutzung prüfen
19.526GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

run-deep-swe

davidondrej

Score any AI model on the DeepSWE coding-agent benchmark via the OpenRouter API. Use when the user wants an independent, reproducible coding-agent eval — "run DeepSWE",…

Preis unbestätigtEntwicklungClaude Code
3850GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

SWE AF

Agent-Field

Autonomous software engineering fleet of AI agents for production-grade PRs on AgentField: plan, code, test, and ship.

Preis unbestätigtEntwicklungClaude CodeOpenAI AgentsVor Nutzung prüfen
969GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

bench

arbazkhan971

Formal benchmark harness. Runs a metric command N times across 2-3 variants (git refs or state-prep shell commands), checks variance, computes delta vs. declared baselin…

Preis unbestätigtEntwicklungClaude Code
26GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

PRarena

aavetis

This repo tracks the opened and merged PRs by the top SWE coding agents by OpenAI, GitHub, and others. Updates regularly.

Preis unbestätigtEntwicklungOpenAI AgentsVor Nutzung prüfen
300GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

skvm-general

SJTU-IPADS

Drive the skvm CLI on behalf of a user to profile models, AOT-compile skills, run skill-assisted tasks, run benchmarks, and manage compiled proposals. Trigger when the u…

Preis unbestätigtEntwicklungClaude Code
552GitHub
Skill ansehen

Anleitungen & Vergleiche

Für Entwickler