No visual example yet
Explore the skillMini Swe Agent
SWE-agent
The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench veri…
OPENAGENTSKILL / DIRECTORY
Temukan skill untuk tugas berikutnya dengan Codex, Claude Code, Cursor, dan lainnya.
14 Skills
Hasil: 14
No visual example yet
Explore the skillSWE-agent
The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench veri…
No visual example yet
Explore the skillJARVIS-Xs
SE-Agent is a self-evolution framework for LLM Code agents. It enables trajectory-level evolution to exchange information across reasoning paths via Revision, Recombinat…
No visual example yet
Explore the skillKodezi
Kodezi Chronos is a debugging-first language model that achieves state-of-the-art results on SWE-bench Lite (80.33%) and 67% real-world fix accuracy, over six times bett…
No visual example yet
Explore the skillchina-qijizhifeng
Official AHE code — Agentic Harness Engineering: observability-driven automatic evolution of coding-agent harnesses (concurrent w/ meta-harness). NexAU-AHE reaches 84.7%…
No visual example yet
Explore the skillalibaba
An Alibaba open-source multi-language benchmark for evaluating LLMs in repository-level automatic code review, featuring an AI-assisted and expert-verified dataset.
No visual example yet
Explore the skillSWE-agent
SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding chal…
No visual example yet
Explore the skilllangchain-ai
An Open-Source Asynchronous Coding Agent
No visual example yet
Explore the skilldavidondrej
Score any AI model on the DeepSWE coding-agent benchmark via the OpenRouter API. Use when the user wants an independent, reproducible coding-agent eval — "run DeepSWE",…
No visual example yet
Explore the skillAgent-Field
Autonomous software engineering fleet of AI agents for production-grade PRs on AgentField: plan, code, test, and ship.
No visual example yet
Explore the skilldatacurve-ai
Measuring frontier coding agents on original, long-horizon engineering tasks
No visual example yet
Explore the skillarbazkhan971
Formal benchmark harness. Runs a metric command N times across 2-3 variants (git refs or state-prep shell commands), checks variance, computes delta vs. declared baselin…
No visual example yet
Explore the skillaavetis
This repo tracks the opened and merged PRs by the top SWE coding agents by OpenAI, GitHub, and others. Updates regularly.
No visual example yet
Explore the skillmicrosoft
A simple SWE style browser agent framework that achieves SOTA results on long horizon web tasks.
No visual example yet
Explore the skillSJTU-IPADS
Drive the skvm CLI on behalf of a user to profile models, AOT-compile skills, run skill-assisted tasks, run benchmarks, and manage compiled proposals. Trigger when the u…