agent-evaluation

Design and implement comprehensive evaluation systems for AI agents. Use when building evals for coding agents, conversational agents, research agents, or computer-use agents. Covers grader types, benchmarks, 8-step roadmap, and production integration.

1.05x

Quality

88%

Does it follow best practices?

Impact

81%

1.05x

Average score across 3 eval scenarios

Securityby

Passed

No known issues

No security issues found

Scanned about 2 months ago

Repository: supercent-io/skills-template
Commit: fd18296

Audited: about 2 months ago
Security analysis

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.