Testing Skill Leaderboard
An agentic skills framework for brainstorming designs, dispatching parallel agents, and standardizing AI-assisted development processes.
A set of AI skills for full-stack development, including Angular 17+ architecture, .NET Core expertise, and system design documentation.
A technical reference for Android development covering CameraX implementation, Jetpack Compose adaptive layouts, and AGP 9 migration.
Security auditing skills for vulnerability hunting, threat modeling, and detecting prompt injection risks in AI-integrated CI/CD pipelines.
A browser automation skill for AI agents to navigate websites, fill forms, scrape data, and test web applications using persistent named pages.
A library of agent skills for .NET development, covering SDK configuration, ASP.NET Core API construction, Blazor components, and test execution.
Technical patterns for Node.js and React development, ClickHouse query optimization, and eval-driven development frameworks.
Provides specialized skills for Go performance profiling, benchmarking, and the development of structured CLI applications.
Browser automation using Playwright for UX validation, responsive checks, and automated web testing.
A collection of Vue 3 skills covering composable design, component architecture, reactivity debugging, Pinia state management, Vue Router patterns, and testing with Vitest and Playwright.
Provides auditing and optimization skills for WCAG 2.2 accessibility, SEO visibility, and Core Web Vitals performance.
A meta-skill for engineering, evaluating, and governing the lifecycle of reusable AI agent skills from workflows and SOPs.
Migrates Apple/Swift Expo native modules from the 1.0 definition DSL to the 2.0 macro API.
Generates and optimizes 3D models, animations, textures, and audio assets specifically for Three.js game development.
Diagnose-first guidance for Terraform/OpenTofu: identifies failure modes, proposes risk-controlled remediation, and enforces a response contract with validation and rollback plans.
Implements a think-act-prove orchestration method for AI agents featuring domain research and adversarial verification of outputs.
A 45-skill kit for Claude Code covering backend development, databases, authentication, devops, multimodal AI, document editing, and debugging.
Over thirty curated skills for marketing lead research, competitive ad analysis, content writing, generative art, frontend development, Notion integration, document processing, and operational workflow automation.
A set of tools and guides for designing, testing, and refining AI skills and agents using context engineering and evaluation metrics.
A set of skills for LLM evaluation, covering synthetic data generation, judge prompt engineering, RAG evaluation, and the creation of human-in-the-loop review interfaces.
Developer skills for AI systems, data pipelines, APIs, architecture, accessibility, security, performance, and framework migration.
A set of agent-optimized skills for automating UI interactions and bug hunting across Apple and Android platforms.
An architecture audit and technical debt assessment tool with automated remediation capabilities for codebase maintenance.
A suite of tools and templates for creating, reviewing, and governing custom Claude Code skills.
A set of skills for creating synthetic data, writing judge prompts, evaluating RAG performance, and building custom review interfaces for error analysis.
A set of developer tools for designing technical specifications and generating AGENTS.md context files to bootstrap AI-assisted development.
Scans Xcode projects for App Store rejection patterns using app-type checklists and asc CLI metadata. Outputs severity-ordered reports with fix suggestions for metadata, subscriptions, privacy, entitlements, and design compliance.
Tools for building Axiom dashboards via API and executing metrics queries against MetricsDB.
Provides 29 scripts for iOS app build automation, simulator lifecycle management, and accessibility-driven UI testing.
Five guard skills for reviewing AI-generated code, documentation, tests, WordPress plugins, and WooCommerce extensions against systematic failure modes.
Generates Solidity fuzz suites for Echidna and Medusa from Foundry or Hardhat projects and converts natural language properties into Solidity assertions.
Technical implementation guides for Akka.NET actor patterns, cluster bootstrapping, and .NET Aspire orchestration.
Captures before-and-after screenshots of web pages or elements for visual comparison and PR documentation.
Technical guidance for Apple platform development covering Swift 6 concurrency, Xcode diagnostics, accessibility auditing, and on-device AI implementation.
Orchestration skills for agent-to-agent handoffs, transcript management, and contract-based behavior validation.
A set of skills for maintaining AGENTS.md and CLAUDE.md instruction files and generating conventional Sentry-style commit messages.
Plan, build and launch products with AI: guided planning prompts, an AI App Starter for Next.js, Supabase and Vercel, and agent configurations.
Optimizes Claude Code skills by running autonomous loops that score outputs against binary evals and mutate prompts to improve reliability.
Evaluates and improves AI agent skills by testing them against golden cases in isolated subagents to drive iterative edits.
A suite of tools for creating Canvas Apps and configuring connections to SharePoint, Dataverse, SQL, and other API connectors.
A set of DevOps skills for building custom AWS and Azure machine images using Packer and managing their registration in HCP Packer.
A set of Cursor skills for implementing Auth.js, PostHog analytics, OpenAPI documentation, and accessibility auditing.
Developer tooling skills for automating GitHub Actions CI fixes, pull request management, and BigQuery analysis artifact generation.
A library of coding agent skills featuring accessibility auditing and active memory management for consistent AI development workflows.
Java engineering skills for architecture review, API contract validation, and refactoring legacy code using clean code principles.
Provides a systematic method for designing, modifying, and reviewing HTTP/JSON API endpoints, request/response shapes, and versioning strategies.
A set of skills for visualizing dbt model lineage via Mermaid diagrams, creating unit tests with mocked inputs, and auditing skill security and quality.
A set of agent workflows for backend technical decision-making, anti-pattern detection, and documentation criteria.
A skill library for bulk code refactoring, test fixing, feature planning, codebase auditing, project bootstrapping, and HTML visual documentation generation.
A set of directorial skills for AI video production focusing on motion language, visual metaphors, and anti-PPT quality control.