CtrlK
BlogDocsLog inGet started
Tessl Logo

backtesting-frameworks

Build robust, production-grade backtesting systems that avoid common pitfalls and produce reliable strategy performance estimates.

44

Quality

45%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/AI-Agents-Safe-Coding-Skills-claude/skills/backtesting-frameworks/SKILL.md

The canonical home for this skill is backtesting-frameworks in rmyndharis/antigravity-skills

SKILL.md
Quality
Evals
Security

Quality

Content

57%Scale 1-3

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The skill is well-structured and concise with good progressive disclosure to external resources. However, it severely lacks actionability - the instructions read as a high-level checklist rather than executable guidance. Without concrete code examples, specific commands, or detailed implementation patterns in the main skill file, Claude would struggle to actually implement a backtesting system.

Suggestions

Add at least one concrete, executable code example showing a minimal backtesting loop structure (e.g., event-driven simulation skeleton)

Include specific validation checkpoints in the workflow, such as 'Verify no lookahead bias by checking data timestamps before each signal'

Provide concrete examples of what 'realistic cost models' and 'point-in-time data pipelines' look like in practice, even if brief

DimensionReasoningScore

Conciseness

The content is lean and efficient, avoiding unnecessary explanations of concepts Claude already knows. Each bullet point earns its place without padding or verbose context.

3 / 3

Actionability

The instructions are vague and abstract with no concrete code, commands, or executable examples. Phrases like 'Build point-in-time data pipelines' and 'Implement event-driven simulation' describe rather than instruct.

1 / 3

Workflow Clarity

Steps are listed in a logical sequence but lack validation checkpoints, feedback loops, or specific criteria for when to proceed. No guidance on how to verify each step was completed correctly.

2 / 3

Progressive Disclosure

Clear overview structure with well-signaled one-level-deep reference to the implementation playbook. Content is appropriately split between overview and detailed resources.

3 / 3

Total

9

/

12

Passed

Description

32%Scale 1-3

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description identifies a clear domain (backtesting) but lacks concrete actions and explicit trigger guidance. It reads more like a marketing tagline than a functional skill description, missing the 'Use when...' clause entirely and relying on abstract qualities rather than specific capabilities.

Suggestions

Add a 'Use when...' clause with trigger terms like 'backtest', 'trading strategy', 'historical simulation', 'strategy performance', or 'quantitative trading'.

Replace abstract qualities with concrete actions such as 'simulate historical trades', 'calculate performance metrics (Sharpe, drawdown)', 'detect lookahead bias', or 'generate equity curves'.

Include file type or domain keywords users might mention like 'OHLCV data', 'price history', 'trading signals', or specific asset classes.

DimensionReasoningScore

Specificity

Names the domain (backtesting systems) and mentions some qualities (robust, production-grade, avoid pitfalls, reliable estimates), but doesn't list concrete actions like 'simulate trades', 'calculate Sharpe ratios', or 'generate equity curves'.

2 / 3

Completeness

Describes what it does (build backtesting systems) but completely lacks a 'Use when...' clause or any explicit trigger guidance for when Claude should select this skill.

1 / 3

Trigger Term Quality

Includes 'backtesting' which is a relevant keyword, but misses common variations users might say like 'backtest', 'strategy testing', 'historical simulation', 'trading strategy', or 'performance analysis'.

2 / 3

Distinctiveness Conflict Risk

'Backtesting systems' is fairly specific to quantitative finance, but 'production-grade' and 'strategy performance' could overlap with general software engineering or analytics skills without clearer domain boundaries.

2 / 3

Total

7

/

12

Passed

Validation

90%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation10 / 11 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

10

/

11

Passed

Repository
administrakt0r/AI-Agents-Safe-Coding-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.