| Skill | Added | Review |
|---|---|---|
evaluate-environments skills/evaluate-environments/SKILL.md Run and evaluate verifiers tasksets. Set up the necessary config files and observe the runs and their results. | 60 60 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: c51c094 | |
create-environments skills/create-environments/SKILL.md Create or migrate native verifiers.v1 taskset, environment, and harness packages. Use to build a taskset, port a benchmark, add task tools, script or model a user, build a multi-agent environment, package an agent harness, or migrate an existing v0 environment to the typed v1 trace model. | 72 72 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: c51c094 | |
brainstorm skills/brainstorm/SKILL.md Run interactive brainstorming across verifiers tasksets, evaluations, GEPA, and RL training. Use when the user wants ideation, literature scanning, concept teaching, roadmap planning, or research program design grounded in local CLI sources, verifiers, and RL trainer code. | 68 68 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: c51c094 |