CtrlK
BlogDocsLog inGet started
Tessl Logo

agents-sdk

Build, run, deploy, and evaluate OpenAI Agents SDK apps from Codex. Use when the user asks to create or adapt an Agents SDK app, build from a prompt or Codex thread, prepare a runnable agent prototype, add a focused eval harness, or deploy locally through the Agents SDK Deployment Manager.

72

Quality

90%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured operational guide with concrete commands, explicit validation checkpoints in the deploy workflow, and a clear done criteria checklist. Its weaknesses are minor: some repeated guidance across sections and the absence of a minimal executable Agent example, with semantics delegated to external docs.

Suggestions

State the PORT/uv-run startup contract once (e.g., in Build Workflow) and reference it from the Deploy checklist instead of repeating the full behavior in both sections.

Add a minimal executable Agent example (a few lines showing `Agent` with static instructions and `Runner.run`) so the build workflow has a copy-paste-ready starting point alongside the delegated docs.

Remove the duplicated doc-gating sentences in the Eval Workflow section since the Rules section already mandates reading the Agents and Agent evals guides.

DimensionReasoningScore

Conciseness

The body is instruction-dense with no padding explaining concepts Claude already knows, but there are minor repetitions: the PORT/`uv run python main.py` startup behavior appears in both Build Workflow and Deploy Workflow, and the doc-reading rules from the Rules section are restated in the Eval Workflow section. These are trimmable instances consistent with anchor 4 rather than the fully lean anchor 5 or the noticeably padded anchor 3.

4 / 5

Actionability

Concrete, copy-paste-ready commands are provided throughout (make deploy invocations, curl health checks, exact layout trees, `git pull --ff-only`), but there is no minimal executable Agent code sample; build semantics are delegated to external docs. This is mostly executable guidance with minor gaps (anchor 4), not fully executable coverage of the common build case (anchor 5).

4 / 5

Workflow Clarity

Build and Deploy workflows are explicitly sequenced with verification steps (manager health, `/health`, session/container checks, `git status --short`), error-recovery guidance ("Stop and report local changes or diverged history instead of forcing the checkout"), and a Done Criteria checklist. This matches the anchor for a clear sequence with explicit validation steps, feedback loops, and checklists.

5 / 5

Progressive Disclosure

No bundle files exist; all references are external URLs listed one level deep in a clearly signaled References section, and sections are well organized. The roughly 170-line body inlines deploy and eval detail that could be split into separate reference files, which is a minor organization gap fitting anchor 4 rather than the cleanly split structure of anchor 5.

4 / 5

Total

17

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that states concrete lifecycle actions (build, run, deploy, evaluate) and follows with an explicit, multi-trigger "Use when" clause covering the natural ways a user would request this skill. It is concise, in third person, and clearly distinguishable from generic app-building skills.

DimensionReasoningScore

Specificity

"Build, run, deploy, and evaluate OpenAI Agents SDK apps from Codex" lists four concrete actions covering the full app lifecycle, matching the anchor for multiple specific concrete actions with comprehensive coverage. It does not fit anchor 4, since there are no minor gaps in the action coverage for this domain.

5 / 5

Completeness

The first sentence explicitly states what the skill does and the second provides an explicit "Use when..." clause with concrete trigger phrases, matching the anchor-5 example structure exactly. It is not anchor 4 because the "when" guidance is fully explicit with multiple concrete triggers rather than only adequately specific.

5 / 5

Trigger Term Quality

The "Use when" clause enumerates natural user phrasings comprehensively: "create or adapt an Agents SDK app", "build from a prompt or Codex thread", "prepare a runnable agent prototype", "add a focused eval harness", and "deploy locally through the Agents SDK Deployment Manager". No common way a user would request this skill is missing, so it exceeds anchor 4's "a few natural terms missing".

5 / 5

Distinctiveness Conflict Risk

The scope is pinned to "OpenAI Agents SDK apps from Codex" and the "Agents SDK Deployment Manager", giving a clear niche with distinct triggers and minimal conflict risk with generic build/deploy skills. It does not overlap meaningfully with adjacent skills, so anchor 4's "minor overlap risk" does not apply.

5 / 5

Total

20

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

15

/

16

Passed

Repository
openai/plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.