github.com/cekura-ai/cekura-skills
| Skill | Added | Review |
|---|---|---|
cekura-metric-design cekura/skills/cekura-metric-design/SKILL.md Use when the user asks to "create a metric", "write a metric", "design a metric", "build a metric for", "evaluate agent performance", "measure call quality", "track a KPI", "add a workflow metric", "improve my metric", "fix a metric", "debug metric results", "set up quality scoring", or "what metrics do I need". Also relevant when discussing LLM judge prompts, custom code metrics, evaluation triggers, VALID_SKIP patterns, section extraction, or metric best practices for Cekura voice AI agents. Covers both creating new metrics and reviewing, iterating on, or troubleshooting existing ones. | 86 86 2.20x Agent success vs baseline Impact 97% 2.20xAverage score across 2 eval scenarios Securityby Low Low-risk findings worth noting Reviewed: Version: af2ceb6 | |
cekura-predefined-metrics cekura/skills/cekura-predefined-metrics/SKILL.md Use when the user asks "what predefined metrics are available", "which built-in metrics should I use", "what does CSAT measure", "how does hallucination detection work", "what's the difference between Interruption Score and AI Interrupting User", "which metrics are free", "which metrics need audio", "configure silence threshold", "set up sentiment metric", or any question about Cekura's out-of-the-box metrics. Covers the full catalog of predefined metrics — what each does, costs, constraints, configuration options, and when to use each one. | 68 68 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: af2ceb6 | |
cekura-self-improving-agent cekura/skills/cekura-self-improving-agent/SKILL.md Use to close the loop on agent quality — turn a failure signal into a verified fix. Triggers: "improve my agent", "self-improving agent", "auto-tune / iterate on my prompt", "fix my agent from test results", "optimize my prompt based on failures", "rewrite my prompt". ALSO for production-call bug fixing: "fix this prod call issue", "debug and fix call ID", "reproduce this production bug". Works across VAPI, Retell, ElevenLabs, Bland, and self-hosted agents, and across three fix surfaces — prompt, tool config, and (self-hosted) owned source code, including infra-flavored / forked-SDK bugs, which are reproduced and validated on Cekura (never a code test). | 70 70 Impact — No eval scenarios have been run Securityby High Do not use without reviewing Version: af2ceb6 | |
cekura-coordinator cekura/skills/cekura-coordinator/SKILL.md Use when the user asks "what can Cekura do", "what commands are available", "help me with Cekura", "what skills do I have", "show me Cekura features", "what's available", "how do I use Cekura", or needs guidance on which Cekura skill to use for their task. Also relevant as the entry point when a user has just installed cekura-skills for the first time. | 63 63 Impact — No eval scenarios have been run Securityby Passed No findings from the security scan Version: af2ceb6 | |
cekura-onboarding cekura/skills/cekura-onboarding/SKILL.md Use when the user says "get started with Cekura", "set up Cekura", "onboard to Cekura", "I'm new to Cekura", "help me set up my agent", "how do I use Cekura", "walk me through Cekura", "configure my project", "first time using Cekura", or needs guidance on initial platform setup. Covers two onboarding paths: **testing** (default — build evaluators and run simulated calls) and **observability** (ingest production call logs and evaluate them). | 65 65 Impact — No eval scenarios have been run Securityby Low Low-risk findings worth noting Version: af2ceb6 | |
cekura-eval-design cekura/skills/cekura-eval-design/SKILL.md Use when the user asks to "create an evaluator", "create evals", "create a scenario", "write a test scenario", "design a test case", "test my agent", "build eval coverage", "plan a test suite", "create red team tests", "set up test profiles", "configure conditional actions", "write a conditional action evaluator", "build a deterministic test", "design an IVR test", "IVR navigation test", "write a unit test for a voice agent", "build a regression test", "scripted scenario", "scripted voice test", "structured evaluator", "exact flow test", "sequential conditions", "fixed sequence test", or "run evals". Also for debugging how the testing agent speaks — "why did it read the number as a word", "make it spell digits", "wrong language" — via scenario_language, personality, and XML tags. Covers evaluator design, coverage strategy, test profiles, mock-tool data, conditional actions (deterministic / unit test / regression / IVR flows), and workflow / red-team / edge-case best practices. | 72 72 Impact — No eval scenarios have been run Securityby Critical Do not install without reviewing Version: af2ceb6 | |
cekura-metric-improvement cekura/skills/cekura-metric-improvement/SKILL.md Use when the user asks to "improve a metric", "run labs", "leave feedback on a metric", "add to labs", "fix metric accuracy", "review metric results", "find misaligned metrics", or "iterate on metric quality". Covers the metric improvement cycle, the feedback workflow, and the labs pipeline used to refine metric accuracy over time. | 65 65 Impact — No eval scenarios have been run Securityby Low Low-risk findings worth noting Version: af2ceb6 | |
cekura-fixing-prod-issues cekura/skills/cekura-fixing-prod-issues/SKILL.md Debugs a failing production call, reproduces the bug with Cekura evaluators, implements a fix, verifies it, runs regression tests, then raises a PR with evidence. Use when the user wants to fix a production call bug, investigate a failing prod call, reproduce and fix a production issue, run regression tests before a PR, or says things like "fix this prod call issue", "debug and fix call ID", "test my fix against prod scenarios", "reproduce this production bug", or "regression test before raising PR". | 69 69 Impact — No eval scenarios have been run Securityby Low Low-risk findings worth noting Version: af2ceb6 | |
cekura-create-agent cekura/skills/cekura-create-agent/SKILL.md Use when the user asks to "create a main agent", "set up a main agent", "add my main agent to Cekura", "configure my main agent", "connect my main agent", "set up mock tools", "add tools to my agent", "upload knowledge base", "configure integration", "connect VAPI", "connect Retell", "connect LiveKit", "connect ElevenLabs", "add dynamic variables", or needs to onboard a voice AI agent onto the Cekura platform. Covers the full agent setup flow: project selection, provider selection, basics and connection type, description, main agent creation, mock tools, knowledge base, dynamic variables, and advanced configuration. | 72 72 Impact — No eval scenarios have been run Securityby Critical Do not install without reviewing Version: af2ceb6 |