What problem does it solve?
Manual transcription of audio content is time-consuming, especially for multilingual recordings, podcasts, or non-English meetings. This Skill automates speech-to-text conversion, cutting transcription work from hours to minutes without manual typing.
Core Features & Use Cases
- Multilingual Transcription: Convert speech in 99 languages to accurate text, including low-resource languages with specialized model support.
- English Translation: Automatically translate non-English audio content to English text for cross-team collaboration.
- Timestamped Output: Generate time-stamped transcripts for subtitles, meeting notes, or audio editing workflows.
- Use Case: A podcast producer can use this Skill to transcribe a 1-hour Spanish-language episode, translate it to English, and export SRT subtitle files for global audiences in under 10 minutes.
Quick Start
Use the whisper skill to transcribe the audio file 'team-standup.mp3' and output a timestamped English transcript.