CtrlK
BlogDocsLog inGet started
Tessl Logo

azure-text-to-speech

Generate neural narration audio using Azure AI Speech (REST text-to-speech). Use when synthesizing voiceovers or narration in OpenMontage. Optional cloud TTS provider — preferred when AZURE_SPEECH_KEY is configured; the local piper_tts remains the default offline path. Shares one Speech resource with azure_stt.

71

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

The canonical home for this skill is azure-text-to-speech in calesthio/OpenMontage

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A lean, highly actionable single-purpose provider skill: executable code, concrete voice/parameter/cost specifics, explicit fallback routing, and genuine batch-verification checkpoints. The only notable inefficiency is the intro paragraph duplicating the frontmatter description.

DimensionReasoningScore

Conciseness

The body is dense and assumes competence — it explains only OpenMontage-specific and Azure-specific facts (fallback chain, cost conventions, determinism), never generic TTS concepts. Not 5 — the opening paragraph restates the frontmatter description nearly verbatim ("Optional cloud TTS provider — ... piper_tts remains the default offline path") and could be trimmed.

4 / 5

Actionability

Copy-paste-ready registry call with real parameter values (voice alias, "-4%" rate, output_path/format), executable env-var setup, a concrete curated voice table, concrete param ranges, and per-call cost math. Not 4 — the common case is fully covered; locale and style are advanced options and are flagged as such.

5 / 5

Workflow Clarity

A single-action skill whose flow is unambiguous and presented in execution order (setup env → availability gate → generate per segment → fallback chain), and batch operations carry explicit validation checkpoints ("listen to the first generated segment before batch-running a full script", "listen to a sample before batch runs"), so the batch cap does not apply. Not 4 — the verification steps are explicit, not merely implied.

5 / 5

Progressive Disclosure

No bundle files exist and none are needed: the curated five-voice shortlist, parameters, and cost notes are appropriately sized for SKILL.md, sections are clearly organized, and the only external links (MS REST docs, voice gallery, pricing) are one level deep and well-signaled. Not 4 — nothing inlined here would be better split into a separate file at this size.

5 / 5

Total

19

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, concrete description that answers both what and when with real trigger phrases and explicitly positions the skill against its sibling TTS providers. Its only weaknesses are a single-action capability list and trigger terms that overlap with sibling TTS skills, mitigated by the AZURE_SPEECH_KEY preference condition.

DimensionReasoningScore

Specificity

"Generate neural narration audio using Azure AI Speech (REST text-to-speech)" names a precise action and mechanism, and "Shares one Speech resource with azure_stt" adds a concrete setup capability. Not 3 — the action is specific rather than generic like "Processes PDF files"; not 5 — only one core action is listed rather than several.

4 / 5

Completeness

Both what ("Generate neural narration audio using Azure AI Speech (REST text-to-speech)") and when ("Use when synthesizing voiceovers or narration in OpenMontage") are explicit, with concrete trigger phrases and an explicit preference condition (AZURE_SPEECH_KEY configured). Not 4 — the when clause is concrete rather than vague.

5 / 5

Trigger Term Quality

Natural trigger terms include "voiceovers", "narration", "text-to-speech", and "TTS", giving good synonym coverage. Not 5 — common phrasings like "voice", "speech", or "audio" as standalone triggers are missing.

4 / 5

Distinctiveness Conflict Risk

It explicitly disambiguates from siblings ("the local piper_tts remains the default offline path", "Shares one Speech resource with azure_stt"), but the "voiceovers or narration" triggers equally match other TTS skills (elevenlabs_tts, openai_tts), leaving minor overlap risk. Not 5 for that reason; not 3 — the differentiation is explicit, not incidental.

4 / 5

Total

17

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

Total

14

/

16

Passed

Repository
calesthio/OpenMontage
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.