audio-transcriber

Transcribe audio recordings into structured Markdown reports with metadata and summaries.

46|14|Updated Jan 11, 2026
One-click install
npx skills add https://github.com/alexdcd/Mafia-Claude-Skills --skill audio-transcriber-alexdcd
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: audio-transcriber
Source: https://github.com/alexdcd/Mafia-Claude-Skills/tree/main/skills/audio-transcriber
Command: npx skills add https://github.com/alexdcd/Mafia-Claude-Skills --skill audio-transcriber-alexdcd

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires faster-whisper, rich, tqdm, whisper, and includes scripts (resource) components.

What problem does it solve?

Transcribes audio to Markdown with automatic metadata extraction and structured summaries, saving time on manual note-taking and distribution.

Core Features & Use Cases

  • Transcribes audio in multiple formats (MP3, WAV, M4A, OGG, FLAC, WEBM) with language detection and speaker labeling.
  • Generates meeting minutes and executive summaries, including actions and decisions.
  • Exports Markdown reports suitable for sharing and archival documentation.
  • Provides subtitle support (SRT/VTT) for video content.

Quick Start

Speak your command: transcribe this meeting audio to Markdown and generate a structured report.

Frequently Asked Questions about audio-transcriber

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I transcribe meeting audio to Markdown with speaker labels?

To transcribe meeting audio to Markdown, the skill uses Faster-Whisper or Whisper to process formats like MP3 and WAV. It features automatic language detection and speaker diarization to generate structured Markdown reports.

Can I generate meeting minutes and executive summaries from audio recordings?

Yes, you can generate meeting minutes and executive summaries from audio recordings. The skill extracts rich metadata and uses optional LLM tools to identify actions and decisions, outputting structured Markdown summaries suitable for sharing.

Does this transcription tool support subtitle generation for video content?

Yes, this transcription tool supports subtitle generation for video content by exporting SRT and VTT files. It processes audio from multiple formats and outputs structured Markdown reports alongside the subtitles.

Do I need Faster-Whisper installed to transcribe audio files?

Yes, you need Faster-Whisper or Whisper installed to transcribe audio files. These dependencies handle the core speech-to-text processing, while optional LLM tools provide intelligent processing for generating summaries and meeting minutes.

What audio formats can I convert to structured notes?

You can convert MP3, WAV, M4A, OGG, FLAC, and WEBM audio formats to structured notes. The skill processes these files to output polished Markdown reports with rich metadata, language detection, and optional summaries.