Leaderboard/Research/autoresearch-skill
Last commit on March 18, 2026·Created on March 18, 2026

olelehmann1337/autoresearch-skill

An autonomous experimentation loop for refining and hardening Claude Code skill prompts.
Combined rank
#300
across all skills
In Research
#19
category rank
Stars
980
7d change unavailable
Forks
117
Watchers
3
Traction scoreGitHub stars can be faked, so popularity alone can be misleading. Traction Score looks for broader signs of real attention, adoption, and active maintenance.
TL;DR

Many AI skills suffer from inconsistent output quality. This tool applies an autoresearch methodology to identify these failures, scoring outputs against binary evaluations and iteratively mutating prompts to eliminate errors.

It helps developers move beyond manual prompt engineering by automating the cycle of testing, scoring, and improving SKILL.md files through repeated execution loops.

WHO IT'S FOR
AI agent / automation builders
optimizing prompt performance via autonomous loops
Prompt engineers
reducing failure rates in complex AI skills
Technical leads / architects
benchmarking and validating AI skill reliability
Platform / DevEx teams
standardizing AI skill quality across repositories
Repository contents

1 skill file

Compatible AgentsThe repository documents support for these agents. The skills may also work with other agents that can load SKILL.md files, but they may need some setup or small changes.
Want to install it?
Manual CopyRepository instructions