CtrlK
BlogDocsLog inGet started
Tessl Logo

audio-transcription

Transcribe local audio/video and Apple Voice Memos quickly with cached MLX Whisper models, including bad/low-quality audio.

64

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/audio-transcription/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

82%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is highly actionable with executable commands and useful feedback loops for hallucination detection, and is well-structured into clear sections; it is slightly above the lean/efficient bar on conciseness and could optionally split the manual template or model cache details into a reference file.

DimensionReasoningScore

Conciseness

The body is mostly lean and command-driven with little explanation of concepts Claude already knows; a few lines (e.g. the 'Fetching 4 files: 100% almost instantly' aside and repeated 'bad/low-quality audio' phrasing) could be trimmed, keeping it just below the fully efficient anchor.

4 / 5

Actionability

Provides fully executable, copy-paste-ready commands covering the common cases — the fast-path invocation, fast/balanced/best/auto variants, the manual mlx_whisper template, precache and cache-verify commands — with concrete flags and output paths.

5 / 5

Workflow Clarity

The fast path describes a clear staged sequence (stable copy, ensure cached model, write outputs, detect hallucination, rerun) and the Quality checks section gives a red-flag checklist with a rerun/cleanup feedback loop; minor gaps remain since the manual path is a single command rather than an explicitly sequenced validated workflow.

4 / 5

Progressive Disclosure

Well-organized into clearly headed sections (Core rules, Fast path, Cached models, Manual command template, Quality checks) in a single self-contained file with no nested references; at ~90 lines it exceeds the under-50-line simple-skill exception, so it sits at 'good structure with minor organization gaps' rather than a top mark.

4 / 5

Total

17

/

20

Passed

Description

70%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and distinctive with strong natural trigger terms, but it omits any explicit 'Use when...' guidance, so the 'when' half of completeness is only weakly implied rather than stated.

Suggestions

Add an explicit trigger clause, e.g. 'Use when the user asks to transcribe an audio or video file, a Voice Memos export, dictation, or bad-quality audio.'

Surface a couple more natural trigger synonyms (dictation, lecture, meeting recording) in the description to broaden keyword coverage.

Drop the filler word 'quickly' — it is a claim rather than a capability and adds no trigger value.

DimensionReasoningScore

Specificity

Names the domain and several concrete actions ('Transcribe local audio/video and Apple Voice Memos', 'cached MLX Whisper models', 'including bad/low-quality audio'), but the core action is essentially one (transcribe) applied to several media types rather than a broad action set, leaving minor coverage gaps.

4 / 5

Completeness

The 'what' is clearly stated (transcribe audio/video and Voice Memos with cached MLX Whisper), but there is no 'Use when...' clause or equivalent explicit trigger guidance in the description, which caps completeness at 3 per the rubric guideline.

3 / 5

Trigger Term Quality

Good natural keyword coverage ('transcribe', 'audio/video', 'Apple Voice Memos', 'bad/low-quality audio') that users would actually say, though common synonyms like 'dictation', 'lecture', or 'meeting recording' (which appear in the body) are absent from the description.

4 / 5

Distinctiveness Conflict Risk

It occupies a clear niche (local MLX Whisper transcription of audio/video and Apple Voice Memos, including bad audio) with distinct triggers and minimal overlap risk against other skills.

5 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
mitsuhiko/agent-stuff
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.