CtrlK
BlogDocsLog inGet started
Tessl Logo

enterprise-agent-ops

Operational controls for long-lived or cloud-hosted agent systems — runtime lifecycle (start, pause, stop, restart), observability (logs, metrics, traces), least-privilege safety scopes and kill switches, and rollout/rollback change management with audit logs and success/cost metrics. Use when running production agent fleets on PM2, systemd, or containers that need monitoring, incident response, or deployment gates.

62

Quality

74%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/enterprise-agent-ops/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

65%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A very lean, well-organized overview that respects the token budget perfectly, but it reads as a policy checklist rather than an operational skill: no commands, thresholds, or concrete procedures anywhere. The incident-response workflow has a sequence and a validation step yet lacks the feedback loop that destructive rollout/rollback operations require.

Suggestions

Add concrete executable commands for the Deployment Integrations section (e.g., pm2 reload / systemctl restart invocations, or a reference file with the actual configurations) so guidance is actionable rather than descriptive.

Extend the Incident Pattern with a feedback loop: what to do when 'run regression + security checks' fails (e.g., keep rollout frozen, revert patch, re-run checks) — required for the destructive rollout/rollback context.

Give at least one concrete instantiation of a Baseline Control or metric (e.g., a sample alert threshold, audit-log entry format, or retry/timeout budget values) so 'hard timeout and retry budgets' and 'success rate' are measurable rather than aspirational.

DimensionReasoningScore

Conciseness

The ~44-line body is entirely lean list items with zero padding and no explanation of concepts Claude already knows — 'immutable deployment artifacts', 'hard timeout and retry budgets', 'cost per successful task' — every token earns its place, matching anchor 5.

5 / 5

Actionability

The body states policy rather than instructing: 'freeze new rollout', 'isolate failing route', 'patch with smallest safe change', 'immutable deployment artifacts' are high-level hints with no commands, thresholds, or concrete procedures (no pm2/systemd commands, no metric thresholds, no example audit-log format), matching anchor 2's 'high-level hints but missing the specific steps to execute'; anchor 3 would require partially concrete, executable guidance.

2 / 5

Workflow Clarity

The Incident Pattern is a clear 6-step sequence with a validation checkpoint ('run regression + security checks' before 'resume gradually'), but rollout/rollback of production fleets is a destructive-change context and there is no feedback loop for what to do when those checks fail — the missing-feedback-loop cap applies, so it cannot exceed anchor 3.

3 / 5

Progressive Disclosure

The skill is under 50 lines with no bundle files (references/, scripts/, assets/ are absent), no inlined content that belongs in a separate file, and well-organized sections (Operational Domains, Baseline Controls, Metrics to Track, Incident Pattern, Deployment Integrations) — the simple-skill exception applies, so well-organized sections alone merit anchor 5.

5 / 5

Total

15

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete, third-person, with an explicit and well-phrased 'Use when' clause naming specific platforms (PM2, systemd, containers) and scenarios (monitoring, incident response, deployment gates). Keyword coverage is good though a few natural synonyms (uptime, health checks, alerts) are absent.

DimensionReasoningScore

Specificity

Enumerates several concrete capability items across four domains — 'runtime lifecycle (start, pause, stop, restart)', 'observability (logs, metrics, traces)', 'least-privilege safety scopes and kill switches', 'rollout/rollback change management with audit logs' — which exceeds anchor 3's 1-2 concrete actions, but stops short of anchor 5's comprehensive coverage since items are domain categories with parenthetical examples and minor gaps remain (e.g., alerting, health checks).

4 / 5

Completeness

Explicitly answers both: what ('Operational controls for long-lived or cloud-hosted agent systems — runtime lifecycle...') and when ('Use when running production agent fleets on PM2, systemd, or containers that need monitoring, incident response, or deployment gates') with concrete trigger phrases in third-person voice, matching anchor 5 exactly.

5 / 5

Trigger Term Quality

Includes natural operator phrases — 'PM2, systemd, or containers', 'monitoring, incident response, or deployment gates', 'kill switches', 'rollback' — matching anchor 4's good keyword coverage, but common synonyms like 'uptime', 'health checks', and 'alerts' are missing, so anchor 5's synonym/extension comprehensiveness is not met.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear niche (production agent fleet operations) with distinct triggers — 'agent fleets', 'PM2', 'systemd', 'kill switches' — that few other skills would claim, but 'containers' and 'deployment gates' leave minor overlap with generic deployment/CI-CD skills, matching anchor 4's 'mostly distinct; minor overlap risk' rather than anchor 5's minimal-conflict bar.

4 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
affaan-m/ECC
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.