github.com/mishatojk/davidskills
Skill | Added | Review |
|---|---|---|
run-deep-swe Score any AI model on the DeepSWE coding-agent benchmark via the OpenRouter API. Use when the user wants an independent, reproducible coding-agent eval — "run DeepSWE", "benchmark this model on DeepSWE", "score model X on the coding benchmark", "test a model via OpenRouter on DeepSWE", or to verify vendor-reported coding scores. Covers setup, the OpenRouter wiring for mini-swe-agent, single-task / subset / full 113-task runs, and leaderboard submission. | 80 80 Impact — No eval scenarios have been run Securityby Advisory Suggest reviewing before use Reviewed: Version: 66860fb | |
teach Teach the user a new skill or concept, within this workspace. | 46 46 Impact — No eval scenarios have been run Securityby Passed No known issues Reviewed: Version: 66860fb | |
pi-custom-model Register a custom or variant model (e.g. an OpenRouter ":nitro" / ":floor" / ":exacto" slug) in the Pi Agent so it can be set as the global default. Use when Pi silently falls back to a different model (e.g. moonshotai/kimi-k2.6) after setting defaultModel, or when a model slug isn't in Pi's bundled list. Triggers on "Pi reset my model", "Pi won't use this model", "add a model to Pi", "Pi default keeps reverting". | 76 76 Impact — No eval scenarios have been run Securityby Passed No known issues Reviewed: Version: 66860fb | |
level-up Gauge the user's technical + product knowledge through 7 adaptive questions, log verbatim answers with honest ratings, and grow a learning plan from the gaps found. Use when the user says "level up", "level-up session", "quiz me", "gauge my knowledge", or wants a new assessment round. Differentiator: this finds and maps gaps; the `teach` skill delivers lessons on them. | 80 80 Impact — No eval scenarios have been run Securityby Passed No known issues Reviewed: Version: 66860fb | |
effective-agent-skills How to write effective agent skills — what to do, what not to do, anatomy, progressive disclosure, design patterns, anti-patterns, testing, security. Read this whenever a skill (Claude Skill, Agent Skill, SKILL.md) is being created, edited, reviewed, or debugged. Use when the user says "create a skill", "new skill", "update this skill", "improve a skill", "why isn't my skill triggering", or anything else involving authoring or editing SKILL.md files. | 68 68 Impact — No eval scenarios have been run Securityby Passed No known issues Reviewed: Version: 66860fb | |
vps-server-management Use when the user wants to manage his VPS servers and the AI agents running inside them — connecting, deploying, monitoring, restarting, and operating remote hosts and their agents. Triggers on VPS, server management, remote host, SSH into server, manage my servers, agents on the server. | 76 76 Impact — No eval scenarios have been run Securityby Passed No known issues Reviewed: Version: 66860fb | |
create-readonly-db-role Provision a hardened SELECT-only Postgres role so AI agents can safely read a production database. Works on Supabase and any Postgres. Use when the user wants agents to query prod data, says "read-only role", "safe prod DB access for agents", or is tired of running SQL by hand for agents. Differentiator: this skill CREATES the role and wiring; day-to-day querying belongs in a project-local skill. | 76 76 Impact — No eval scenarios have been run Securityby Passed No known issues Reviewed: Version: 66860fb | |
codex-subagent Launch OpenAI Codex CLI as a subagent (ChatGPT subscription auth, no API key). Use when delegating a self-contained coding task to Codex from another agent — parallel implementation work, a second opinion, or an independent verification pass. | 75 75 Impact — No eval scenarios have been run Securityby Passed No known issues Reviewed: Version: 66860fb | |
agent-self-scheduling Make an AI agent run on a schedule, loop, or interval — cron, heartbeats, recurring autonomous checks. Use for "run every N minutes", "schedule a task", "run on a loop", "heartbeat". Covers external clocks (Claude Code, Codex, Pi) vs Hermes' built-in scheduler. | 77 77 Impact — No eval scenarios have been run Securityby Advisory Suggest reviewing before use Reviewed: Version: 66860fb |