No visual example yet
Explore the skillMini Swe Agent
SWE-agent
The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench veri…
OPENAGENTSKILL / DIRECTORY
Finde den passenden Skill für deine nächste Aufgabe mit Codex, Claude Code, Cursor und mehr.
36 Skills
Ergebnisse: 36
No visual example yet
Explore the skillSWE-agent
The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench veri…
No visual example yet
Explore the skillJARVIS-Xs
SE-Agent is a self-evolution framework for LLM Code agents. It enables trajectory-level evolution to exchange information across reasoning paths via Revision, Recombinat…
No visual example yet
Explore the skillKodezi
Kodezi Chronos is a debugging-first language model that achieves state-of-the-art results on SWE-bench Lite (80.33%) and 67% real-world fix accuracy, over six times bett…
No visual example yet
Explore the skillaquasecurity
Checks whether Kubernetes is deployed according to security best practices as defined in the CIS Kubernetes Benchmark
No visual example yet
Explore the skillAyanami0730
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
No visual example yet
Explore the skillchina-qijizhifeng
Official AHE code — Agentic Harness Engineering: observability-driven automatic evolution of coding-agent harnesses (concurrent w/ meta-harness). NexAU-AHE reaches 84.7%…
No visual example yet
Explore the skillonyx-dot-app
Dataset and benchmark for RAG on company internal documents.
No visual example yet
Explore the skillalibaba
An Alibaba open-source multi-language benchmark for evaluating LLMs in repository-level automatic code review, featuring an AI-assisted and expert-verified dataset.
No visual example yet
Explore the skillzhikunqingtao
Claude Code / Cursor 开源替代。部署在你自己的服务器上,团队用浏览器打开就能编程——包括手机。CLI & Web UI 双入口,Multi-Agent 协作,原生直连千问/DeepSeek 等国产大模型。SWE-bench Lite 56%,不是玩具。技能/插件/跨会话记忆,8 层安全沙箱,数据不离开你的机器。Doc…
No visual example yet
Explore the skillSWE-agent
SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding chal…
No visual example yet
Explore the skillAndrejOrsula
Robot Learning Beyond Earth
No visual example yet
Explore the skilllangchain-ai
An Open-Source Asynchronous Coding Agent
No visual example yet
Explore the skilldavidondrej
Score any AI model on the DeepSWE coding-agent benchmark via the OpenRouter API. Use when the user wants an independent, reproducible coding-agent eval — "run DeepSWE",…
No visual example yet
Explore the skillAgent-Field
Autonomous software engineering fleet of AI agents for production-grade PRs on AgentField: plan, code, test, and ship.
No visual example yet
Explore the skilldatacurve-ai
Measuring frontier coding agents on original, long-horizon engineering tasks
No visual example yet
Explore the skillarthur-ai