github.com/coder/agent-tty
| Skill | Added | Review |
|---|---|---|
agent-terminal dogfood/20260327-public-skill/installed-skill/SKILL.md Terminal and TUI automation CLI for AI agents. Use when the user needs to create a terminal session, run a command in a terminal, automate an interactive CLI or TUI, wait for terminal output, capture a TUI screenshot, export a terminal recording, or test a CLI workflow with reviewable artifacts. | 76 76 Impact — No eval scenarios have been run Securityby — The risk profile of this skill Version: ebff2c2 | |
agent-tty skill-data/agent-tty/SKILL.md Terminal and TUI automation CLI for AI agents. Use when the user needs to create a terminal session, run a command in a terminal, automate an interactive CLI or TUI, wait for terminal output, capture a TUI screenshot, export a terminal recording, or test a CLI workflow with reviewable artifacts. | — | |
agent-tty skills/agent-tty/SKILL.md Terminal and TUI automation CLI for AI agents. Use when the user needs to create a terminal session, run a command in a terminal, automate an interactive CLI or TUI, wait for terminal output, capture a TUI screenshot, export a terminal recording, or test a CLI workflow with reviewable artifacts. | — | |
codebase-design .agents/skills/codebase-design/SKILL.md Shared vocabulary for designing deep modules. Use when the user wants to design or improve a module's interface, find deepening opportunities, decide where a seam goes, make code more testable or AI-navigable, or when another skill needs the deep-module vocabulary. | 72 72 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: ebff2c2 | |
diagnose .agents/skills/diagnose/SKILL.md Disciplined diagnosis loop for hard bugs and performance regressions. Reproduce → minimise → hypothesise → instrument → fix → regression-test. Use when user says "diagnose this" / "debug this", reports a bug, says something is broken/throwing/failing, or describes a performance regression. | 95 95 0.97x Agent success vs baseline Impact 90% 0.97xAverage score across 3 eval scenarios Securityby — The risk profile of this skill Reviewed: Version: ebff2c2 | |
diagnosing-bugs .agents/skills/diagnosing-bugs/SKILL.md Diagnosis loop for hard bugs and performance regressions. Use when the user says "diagnose"/"debug this", or reports something broken/throwing/failing/slow. | 74 74 Impact — No eval scenarios have been run Securityby High Do not use without reviewing Version: ebff2c2 | |
dogfood-tui skill-data/dogfood-tui/SKILL.md Structured TUI dogfooding and QA workflow using agent-tty. Use for exploratory testing, bug hunting, release-readiness validation, and UX review of terminal applications. | — | |
domain-modeling .agents/skills/domain-modeling/SKILL.md Build and sharpen a project's domain model. Use when the user wants to pin down domain terminology or a ubiquitous language, record an architectural decision, or when another skill needs to maintain the domain model. | 60 60 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: ebff2c2 | |
eval-guide .mux/skills/eval-guide/SKILL.md Guide for running statistically meaningful agent-tty evals with trials, parallelism, and A/B comparison. Covers non-determinism baseline, recommended sample sizes, and result interpretation. | 53 53 Impact — No eval scenarios have been run Securityby — The risk profile of this skill Version: ebff2c2 | |
grilling .agents/skills/grilling/SKILL.md Interview the user relentlessly about a plan or design. Use when the user wants to stress-test a plan before building, or uses any 'grill' trigger phrases. | 67 67 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: ebff2c2 | |
grill-with-docs .agents/skills/grill-with-docs/SKILL.md A relentless interview to sharpen a plan or design, which also creates docs (ADR's and glossary) as we go. | 33 33 Impact — No eval scenarios have been run Securityby — The risk profile of this skill Version: ebff2c2 | |
improve .agents/skills/improve/SKILL.md Survey any codebase as a senior advisor and produce prioritized, self-contained implementation plans for OTHER models/agents to execute. Strictly read-only on source code — never implements, fixes, or refactors anything itself. Use when asked to audit a codebase, find improvement opportunities (bugs, security, performance, test coverage, tech debt, migrations, DX), suggest features or where to take the project next (roadmap, product direction), or generate handoff plans for another agent to implement. | 76 76 Impact — No eval scenarios have been run Securityby Low Low-risk findings worth noting Version: ebff2c2 | |
improve-codebase-architecture .agents/skills/improve-codebase-architecture/SKILL.md Scan a codebase for deepening opportunities, present them as a visual HTML report, then grill through whichever one you pick. | 36 36 Impact — No eval scenarios have been run Securityby — The risk profile of this skill Version: ebff2c2 | |
release-maintainer .agents/skills/release-maintainer/SKILL.md Internal maintainer SOP for version bumps, release PRs, tagging, publishing, and post-publish verification in this repository. | 48 48 Impact — No eval scenarios have been run Securityby — The risk profile of this skill Version: ebff2c2 | |
tdd .agents/skills/tdd/SKILL.md Test-driven development. Use when the user wants to build features or fix bugs test-first, mentions "red-green-refactor", or wants integration tests. | 62 62 Impact — No eval scenarios have been run Securityby — The risk profile of this skill Version: ebff2c2 | |
to-issues .agents/skills/to-issues/SKILL.md Break a plan, spec, or PRD into independently-grabbable issues on the project issue tracker using tracer-bullet vertical slices. | 66 66 0.95x Agent success vs baseline Impact 80% 0.95xAverage score across 1 eval scenario Securityby — The risk profile of this skill Reviewed: Version: ebff2c2 | |
to-prd .agents/skills/to-prd/SKILL.md Turn the current conversation into a PRD and publish it to the project issue tracker — no interview, just synthesis of what you've already discussed. | 54 54 0.00x Agent success vs baseline Impact 0% 0.00xAverage score across 1 eval scenario Securityby — The risk profile of this skill Reviewed: Version: ebff2c2 | |
triage .agents/skills/triage/SKILL.md Move issues and external PRs through a state machine of triage roles — categorise, verify, grill if needed, and write agent-ready briefs. | 93 93 1.42x Agent success vs baseline Impact 100% 1.42xAverage score across 2 eval scenarios Securityby — The risk profile of this skill Reviewed: Version: ebff2c2 |