CtrlK
BlogDocsLog inGet started
Tessl Logo

test-reasoning

Validate that reasoning parameters are correctly serialized and sent to provider APIs. Use when the user asks to test reasoning serialization, run reasoning tests, verify reasoning config fields, or check that ReasoningConfig maps correctly to provider-specific JSON (OpenRouter, Anthropic, GitHub Copilot, Codex).

80

Quality

100%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-3

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is concise, fully actionable, and well-structured, with a clear quick-start workflow and one-level references to a real bundled script. No dimension shows notable weakness.

DimensionReasoningScore

Conciseness

Lean body with no padding and no explanation of concepts Claude already knows; every section (Quick Start, manual run, coverage table, references) earns its place.

3 / 3

Actionability

Provides copy-paste-ready executable guidance: './scripts/test-reasoning.sh' and a concrete manual invocation with real env vars plus an explicit output file to inspect.

3 / 3

Workflow Clarity

Quick Start sequences build → run → capture request body → assert JSON fields, and the bundled script performs validation internally; for this single-task skill the workflow is unambiguous. Not a destructive/batch operation, so no validation-loop cap applies.

3 / 3

Progressive Disclosure

Well-organized sections with one-level references: a real bundled script (./scripts/test-reasoning.sh, present in scripts/) and external doc URLs, no deep nesting.

3 / 3

Total

12

/

12

Passed

Description

100%Weight 40%Scale 1-3

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, trigger-rich, and complete, clearly stating both the capability and the conditions under which to invoke it. It carves out a distinct niche with minimal conflict risk.

DimensionReasoningScore

Specificity

Names concrete actions ('Validate that reasoning parameters are correctly serialized and sent to provider APIs', 'check that ReasoningConfig maps correctly to provider-specific JSON') across four named providers, going beyond a single domain mention.

3 / 3

Completeness

Explicitly answers both what ('Validate that reasoning parameters are correctly serialized and sent to provider APIs') and when ('Use when the user asks to test reasoning serialization, run reasoning tests...').

3 / 3

Trigger Term Quality

Includes natural phrases a user would actually say — 'test reasoning serialization', 'run reasoning tests', 'verify reasoning config fields' — with good coverage of variations.

3 / 3

Distinctiveness Conflict Risk

A narrow, specific niche (reasoning serialization for OpenRouter, Anthropic, GitHub Copilot, Codex) with distinct triggers unlikely to overlap with unrelated skills.

3 / 3

Total

12

/

12

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
tailcallhq/forgecode
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.