Skill | Added | Review |
|---|---|---|
ai-regression-testing skills/ai-regression-testing/SKILL.md Regression testing strategies for AI-assisted development. Sandbox-mode API testing without database dependencies, automated bug-check workflows, and patterns to catch AI blind spots where the same model writes and reviews code. | 66 66 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd | |
ai-first-engineering skills/ai-first-engineering/SKILL.md Engineering operating model for teams where AI agents generate a large share of implementation output. | 54 54 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd | |
agentic-os skills/agentic-os/SKILL.md Build persistent multi-agent operating systems on Claude Code. Covers kernel architecture, specialist agents, slash commands, file-based memory, scheduled automation, and state management without external databases. | 61 61 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd | |
agent-self-evaluation skills/agent-self-evaluation/SKILL.md Use after completing any non-trivial task. The agent self-rates its output on 5 axes — accuracy, completeness, clarity, actionability, conciseness — with concrete evidence per criterion. Produces a structured 1-5 scorecard with specific improvement suggestions. | 64 64 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd | |
agent-payment-x402 skills/agent-payment-x402/SKILL.md Add x402 payment execution to AI agents with per-task budgets, spending controls, and non-custodial wallets. Supports Base through agentwallet-sdk and X Layer through OKX Payments / OKX Agent Payments Protocol. | 61 61 Impact — No eval scenarios have been run Securityby Medium Suggest reviewing before use Reviewed: Version: 754b8dd | |
agent-harness-construction skills/agent-harness-construction/SKILL.md Design and optimize AI agent action spaces, tool definitions, and observation formatting for higher completion rates. | 60 60 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd | |
agent-eval skills/agent-eval/SKILL.md Head-to-head comparison of coding agents (Claude Code, Aider, Codex, etc.) on custom tasks with pass rate, cost, time, and consistency metrics | 58 58 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd | |
agent-architecture-audit skills/agent-architecture-audit/SKILL.md Full-stack diagnostic for agent and LLM applications. Audits the 12-layer agent stack for wrapper regression, memory pollution, tool discipline failures, hidden repair loops, and rendering corruption. Produces severity-ranked findings with code-first fixes. Essential for developers building agent applications, autonomous loops, or any LLM-powered feature. | 59 59 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd | |
accessibility skills/accessibility/SKILL.md Design, implement, and audit inclusive digital products using WCAG 2.2 Level AA standards. Use this skill to generate semantic ARIA for Web and accessibility traits for Web and Native platforms (iOS/Android). | 62 62 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Reviewed: Version: 754b8dd |