Audio expert. ALWAYS invoke this skill when the user asks to transcribe, recognize, or convert speech/audio to text.
77
96%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Passed
No findings from the security scan
Use these tools to transcribe audio files to text using speech-to-text models.
Transcribe an audio file to text.
tai tool audio_transcribe --audio_path /path/to/meeting.m4atai tool audio_transcribe --audio_path /path/to/recording.wav --language en --provider llm.my-openai:whisper-1| Parameter | Type | Required | Description |
|---|---|---|---|
| audio_path | string | yes | Audio file path. Supported: mp3, m4a, wav, webm, mp4, mpeg, mpga |
| language | string | no | ISO 639-1 language code (e.g. en, zh, ja). Auto-detected if omitted |
| provider | string | no | STT provider connector ID. If omitted, uses the default STT provider |
List available speech-to-text providers and models.
tai tool audio_providers| Parameter | Type | Required | Description |
|---|---|---|---|
| capability | string | no | Filter by capability (default: audio) |
Returns a list of providers with their available models and connector IDs that can be passed to audio_transcribe.
Only use the parameters listed above for each tool. Do not pass unsupported parameters — they will be ignored or cause errors.
d90d41d
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.