What problem does it solve?
This Skill helps media teams turn recorded or live speech into accurate, reviewable transcripts and timecoded production assets without improvising provider configuration, lifecycle management, privacy controls, or quality assurance.
Core Features & Use Cases
- Transcription Surface Selection: Choose pre-recorded, synchronous short-file, or real-time streaming transcription based on duration, latency, and production needs.
- Production Transcript Workflows: Configure diarization, speaker identification, language detection, code-switching, keyterms, captions, subtitles, timestamps, summaries, chapters, entities, sentiment, translation, profanity filtering, and PII redaction.
- Operational Guardrails: Plan webhooks, rate limits, retries, billing, retention, consent, API-key security, data controls, and human review for production delivery.
- Use Case: Prepare a podcast package with a speaker-reviewed transcript, SRT and VTT caption drafts, chapter markers, pull-quote candidates, and privacy-aware derivative artifacts.
Quick Start
Use the AssemblyAI transcription skill to create a production plan for the attached interview, including the appropriate API surface, speaker strategy, caption outputs, privacy controls, webhook lifecycle, and final QA checks.