What problem does it solve? Audio and video recordings are opaque to an AI assistant until converted to text. This Skill turns speech from recordings, podcasts, meetings, voice memos, lectures, and interviews into timestamped transcripts that can be quoted, summarized, and searched. ## Core Features & Use Cases - Speech-to-Text with Timestamps: Returns transcripts as [MM:SS] lines so you can quote and link specific moments in a recording. - Flexible Input Sources: Accepts conversation asset ids, direct http(s) media URLs, or sandbox file paths, with automatic language detection or an ISO-639-1 hint. - Automatic Long-File Handling: Splits recordings over the ~25 MB provider limit and stitches parts back with absolute timestamps on the full timeline. - Use Case: A user attaches a 45-minute meeting recording and asks for a summary with key quotes. The Skill transcribes the whole file, then the assistant delivers the summary in the reply and the full transcript as an exported file. ## Quick Start Transcribe the attached meeting recording and give me a summary with the key quotes and their timestamps.