songsee

Generate spectrograms and audio feature visualizations from audio files via CLI.

Updated Aug 21, 2026
One-click install
npx skills add https://github.com/TylerSimons1127/vibe --skill songsee-tylersimons1127
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: songsee
Source: https://github.com/TylerSimons1127/vibe/tree/main/skills/media/songsee
Command: npx skills add https://github.com/TylerSimons1127/vibe --skill songsee-tylersimons1127

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires songsee, ffmpeg.

What problem does it solve? Analyzing audio content visually requires specialized tooling; this Skill turns audio files into spectrograms and multi-panel feature visualizations (mel, chroma, MFCC, and more) with a single CLI command, making audio structure, tempo, and onsets inspectable as images. ## Core Features & Use Cases - Spectrogram Generation: Render standard or mel-scaled spectrograms from WAV and MP3 files, with optional ffmpeg support for other formats. - Multi-Panel Feature Grids: Combine up to nine visualization types (chroma, HPSS, self-similarity, loudness, tempogram, MFCC, spectral flux) into a single image. - Time Slicing & Styling: Extract specific time ranges with --start/--duration and customize output with color palettes, dimensions, and frequency filters. - Use Case: Compare two synthesized audio outputs by generating side-by-side mel spectrograms, then inspect the images with vision_analyze to verify the synthesis pipeline. ## Quick Start Ask the agent to generate a mel spectrogram and chroma visualization grid from your audio file, saved as a PNG image.

Frequently Asked Questions about songsee

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate a spectrogram from an MP3 file?▼

Run songsee with the audio file path, for example songsee track.mp3, which writes a spectrogram image by default. Use -o to set the output path and --format to choose png or jpg.

What audio visualization types does songsee support?▼

songsee supports nine types: spectrogram, mel, chroma, hpss, selfsim, loudness, tempogram, mfcc, and flux. Pass them comma-separated via --viz to render multiple panels as a grid in one image.

Does songsee support audio formats other than WAV and MP3?▼

WAV and MP3 are decoded natively by songsee. Other formats require ffmpeg to be installed on the system for decoding.

How do I visualize only part of an audio file?▼

Use the --start and --duration flags to slice the audio, for example songsee track.mp3 --start 12.5 --duration 8 renders an 8-second segment starting at 12.5 seconds.

Can songsee read audio from stdin?▼

Yes, pipe audio into songsee using a dash as the input, such as cat track.mp3 | songsee - --format png -o out.png. This is useful in shell pipelines without writing intermediate files.