qwen3-asr-assistant

Transcribe spoken audio with Qwen3-ASR and rewrite text into emails, notes, or social posts.

5.0k|479|Updated Feb 2, 2026
One-click install
npx skills add https://github.com/anbeime/skill --skill qwen3-asr-assistant
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: qwen3-asr-assistant
Source: https://github.com/anbeime/skill/tree/main/skills/qwen3-asr-assistant/qwen3-asr-assistant
Command: npx skills add https://github.com/anbeime/skill --skill qwen3-asr-assistant

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires requests, numpy, and includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill automates the process of converting spoken audio into text and then intelligently rewriting that text into various formats like emails, notes, or social media posts, streamlining communication and content creation.

Core Features & Use Cases

  • Speech-to-Text: Accurately transcribes audio recordings into written text using the Qwen3-ASR model.
  • Smart Text Rewriting: Transforms transcribed text into polished emails, structured notes, or engaging social media content.
  • Audio Concatenation: Seamlessly merges multiple audio recordings into a single, coherent text document.
  • Use Case: A project manager can record meeting notes, have them transcribed, and then instantly convert the raw text into a formal meeting minutes email to be shared with the team.

Quick Start

Use the qwen3-asr-assistant skill to transcribe the audio file 'meeting_recording.wav' and rewrite the content as a formal email.

Frequently Asked Questions about qwen3-asr-assistant

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert speech to text from audio recordings?

To convert speech to text, the Skill uses the Qwen3-ASR model to accurately transcribe audio recordings into written text. It supports real-time voice recognition and multi-segment audio splicing for comprehensive transcriptions.

Can I automatically rewrite transcribed audio into formatted emails or notes?

Yes, you can automatically rewrite transcribed audio into formatted emails or notes using the smart text rewriting feature. It transforms raw transcribed text into polished emails, structured notes, or engaging social media posts.

How do I transcribe and merge multi-segment audio recordings into a single document?

To transcribe and merge multi-segment audio recordings, the Skill provides audio concatenation capabilities. It seamlessly merges multiple audio recordings into a single, coherent text document for streamlined processing.

Does this speech to text tool require Python dependencies like requests and numpy?

Yes, this speech to text tool requires the Python dependencies requests and numpy to function. These libraries are necessary for handling data processing and API interactions during the audio transcription process.

What is the best way to generate meeting minutes from voice memos?

The best way to generate meeting minutes from voice memos is recording the audio, transcribing it with Qwen3-ASR, and using the rewriting feature to instantly convert raw text into a formal meeting minutes email.

Are there limitations when processing audio files for content creation?

Limitations when processing audio files for content creation depend on the Qwen3-ASR model's accuracy with varying audio quality and languages. The Skill is designed for clear recordings to effectively generate social media posts and notes.