What problem does it solve?
Whisper removes the manual burden of turning spoken audio into searchable, readable text, while also handling translation for multilingual recordings and noisy sources.
Core Features & Use Cases
- Speech-to-Text Transcription: Convert podcasts, meetings, interviews, lectures, and videos into text across 99 languages.
- Translation to English: Translate non-English speech into English for faster review and downstream reuse.
- Operational Flexibility: Choose model sizes from tiny to large or turbo, add timestamps, and process single files or batches for different accuracy and speed needs.
- Use Case: A content team can transcribe an episode, generate subtitles, and create a readable transcript for publishing, search, and accessibility.
Quick Start
Use the whisper skill to transcribe the attached audio file and, if needed, translate it to English.