What problem does it solve?
Audio analysis typically requires specialized tools to convert raw sound files into visual representations of frequency and feature data, a process that is time-consuming and inaccessible without dedicated software. This skill automates that workflow using the songsee CLI, eliminating manual setup and tooling overhead.
Core Features & Use Cases
- Spectrogram Generation: Create standard frequency spectrograms from common audio formats like WAV and MP3, with support for additional formats via ffmpeg.
- Multi-Feature Panel Visualization: Render multiple audio features (mel spectrograms, chroma vectors, MFCCs, loudness, and more) in a single customizable grid for comprehensive analysis.
- Time-Sliced Analysis: Extract and visualize specific segments of long audio files for targeted inspection of particular sections. Use cases include audio engineers analyzing song structure, music researchers comparing feature patterns across tracks, and sound designers identifying frequency anomalies in audio assets.
Quick Start
Use the songsee skill to generate a full multi-panel feature visualization from your audio file 'track.mp3'.