Workflow 1.5: Bridge between idea discovery and auto review. Reads EXPERIMENT_PLAN.md, implements experiment code, deploys to GPU, collects initial results. Use when user says "实现实验", "implement experiments", "bridge", "从计划到跑实验", "deploy the plan", or has an experiment plan ready to execute.
68
83%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Critical
Do not install without reviewing
Security
3 findings: 2 critical severity, 1 high severity. Installing this skill is not recommended: please review these findings carefully if you do intend to do so.
Detected a suspicious URL in the skill instructions that could lead the agent to download and execute malicious scripts or binaries. This includes links to executables from untrusted sources, typosquatting of official packages, URL shorteners that obscure the destination, and personal file hosting services.
This is an untrusted GitHub repository URL (https://github.com/org/project) — cloning and running code from an arbitrary repo can execute arbitrary scripts and binaries, so it could be used to distribute malware unless the repo and its maintainers are verified.
Detected high-risk code patterns in the skill content — including its prompts, tool definitions, and resources — such as data exfiltration, backdoors, remote code execution, credential theft, system compromise, supply chain attacks, and obfuscation techniques.
The skill contains deliberate instructions that can exfiltrate potentially sensitive code/data to an external reviewer by default and perform silent filesystem writes without user consent, posing high risk of data leakage and unauthorized modification.
The skill handles credentials insecurely by requiring the agent to include secret values verbatim in its generated output. This exposes credentials in the agent’s context and conversation history, creating a risk of data exfiltration.
The prompt instructs the agent to "paste the experiment scripts" and other file contents into a spawned-agent message (and asks reviewers to verify setups including "secrets"), which forces any secrets present in those files to be copied verbatim into the LLM-generated output, enabling exfiltration.
Low
Low-risk findings.
1 low severity finding. Worth noting, but not necessarily harmful.
The skill fetches instructions or code from an external URL at runtime, and the fetched content directly controls the agent’s prompts or executes code. This dynamic dependency allows the external source to modify the agent’s behavior without any changes to the skill itself.
The skill clones a remote repository at runtime (git clone <BASE_REPO>) and the example override points to https://github.com/org/project, meaning fetched code from that URL could be executed as experiment scripts and thus directly control runtime behavior.
f4f20f9
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.