What problem does it solve?
This Skill eliminates the tedious, error-prone manual work of transcribing audio recordings, saving hours of effort for teams and individuals working with spoken content like meetings, interviews, and voice notes.
Core Features & Use Cases
- Multi-format Audio Transcription: Supports WAV, MP3, M4A, FLAC, and OGG audio files, plus base64 encoded audio input for flexible integration.
- Batch Processing: Transcribe multiple audio files in a single run, ideal for processing large volumes of meeting recordings or interview files.
- Use Case Example: A content team can use this Skill to transcribe 10 podcast episodes into searchable text in minutes, instead of spending hours manually typing out content.
Quick Start
Use the ASR skill to transcribe the audio file 'team-meeting-recording.wav' into accurate editable text.