CtrlK
BlogDocsLog inGet started
Tessl Logo

do-and-judge

Execute a task with sub-agent implementation and LLM-as-a-judge verification with automatic retry loop

53

Quality

60%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Critical

Do not install without reviewing

Fix and improve this skill with Tessl

tessl review fix ./plugins/sadd/skills/do-and-judge/SKILL.md
SKILL.md
Quality
Evals
Security

Security

1 critical severity finding. Installing this skill is not recommended: please review these findings carefully if you do intend to do so.

Critical

E006: Malicious code pattern detected in skill scripts.

What this means

Detected high-risk code patterns in the skill content — including its prompts, tool definitions, and resources — such as data exfiltration, backdoors, remote code execution, credential theft, system compromise, supply chain attacks, and obfuscation techniques.

Why it was flagged

The skill instructs orchestrators to include an environment variable (CLAUDE_PLUGIN_ROOT) verbatim in prompts to external sub-agents, which intentionally exposes internal environment data to downstream models—this is deliberate credential/exfiltration risk.

Report incorrect finding

Low

Low-risk findings.

1 low severity finding. Worth noting, but not necessarily harmful.

Low

W011: Third-party content exposure detected (indirect prompt injection risk).

What this means

The skill exposes the agent to untrusted, user-generated content from public third-party sources, creating a risk of indirect prompt injection. This includes browsing arbitrary URLs, reading social media posts or forum comments, and analyzing content from unknown websites.

Why it was flagged

Skill.md’s required runtime workflow takes arbitrary free-form text from the user as “Task” and injects it directly into the meta-judge prompt and judge prompt (and also into implementation sub-agent prompts), so an outsider can supply poison via the `task` argument that the LLM ingests before selecting any specific item.

Repository
NeoLabHQ/context-engineering-kit
Audited
Security analysis
Snyk

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.