CtrlK
BlogDocsLog inGet started
Tessl Logo

sentry-backend-bugs

Review Sentry Python and Django changes for bug patterns drawn from real production issues. Use when reviewing a backend diff or PR, checking Warden findings, auditing the current branch, reviewing production-error patterns, or looking for common regressions in `src/` and `tests/`.

93

0.97x
Quality

91%

Does it follow best practices?

Impact

94%

0.97x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered review skill: concrete, data-backed pattern checks with red flags, safe patterns, and explicit do-not-flag exceptions that prevent false positives, plus a routing table to complementary per-category references. Its only real weakness is token weight — motivational statistics and inline check detail that partially duplicates the reference files could be slimmed or pushed into the references.

DimensionReasoningScore

Conciseness

The body is dense with information Claude does not already know — Sentry-specific function names, exceptions, and bug classes — with no basic-concept padding. The corpus statistics line ('638 real production issues (393 resolved, 220 unresolved, 25 ignored)...') and per-check issue/event counts are motivational context rather than instruction and could be trimmed, matching the 4 anchor rather than 5.

4 / 5

Actionability

Fully concrete and executable: specific functions ('by_qualified_short_id_bulk()', 'resolve_apdex_function'), specific exceptions (try/except 'DoesNotExist', 'ApiError', 'IntegrityError'), a copy-ready fix ('min(value, 2_147_483_647)'), explicit HTTP status rules, and a hard rule that fixes must include actual code.

5 / 5

Workflow Clarity

Clear three-step sequence (classify → check patterns → report) with explicit validation checkpoints: a HIGH/MEDIUM/LOW confidence gate table, 'Read the endpoint's parent class before reporting', 'Only report if you can trace a specific input that triggers the bug', and an explicit stop condition ('report zero findings') that prevents issue invention.

5 / 5

Progressive Disclosure

The Step 1 routing table signals eight real, one-level-deep reference files whose contents (real examples, root causes, fix patterns) complement the body. However, ~200 lines of per-check red-flag/safe-pattern detail live inline in SKILL.md and partially overlap the category references (e.g., Check 2 vs missing-records.md), so the split is good but not maximally clean.

4 / 5

Total

18

/

20

Passed

Description

95%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete third-person capability statement paired with an explicit, comprehensive 'Use when...' clause whose triggers (diff/PR/branch review, Warden findings, production-error patterns, regressions) are natural user phrasings. The only minor gap is that it advertises a single capability rather than enumerating several distinct actions.

DimensionReasoningScore

Specificity

Names the domain concretely ('Sentry Python and Django changes for bug patterns drawn from real production issues') and scopes review targets ('backend diff or PR', 'Warden findings', 'current branch', 'src/ and tests/'), but the core capability is a single action — review for bug patterns — rather than the multiple distinct actions of the 5 anchor.

4 / 5

Completeness

Explicitly answers both: sentence one states what it does ('Review Sentry Python and Django changes for bug patterns...') and sentence two gives explicit 'Use when...' triggers covering diffs, PRs, Warden findings, branch audits, and regression checks.

5 / 5

Trigger Term Quality

Comprehensive natural terms with synonyms (review/audit/check, diff/PR/branch) and concrete tokens users would actually say ('Warden findings', 'production-error patterns', 'common regressions', 'src/', 'tests/') — a user asking to review a backend PR or check Warden findings would naturally hit these.

5 / 5

Distinctiveness Conflict Risk

The 'Sentry Python and Django' and 'Warden findings' scoping carves a clear niche with distinct trigger phrases; a general code-review skill would not collide with these terms, so conflict risk is minimal.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
getsentry/sentry
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.