Skip to content
Development
Skill

/prismer-songsee

Audio spectrograms/features (mel, chroma, MFCC) via CLI.

BOOST
From plugin
prismercloud
1.6k102 skills
Install
$ npx -y skills add Prismer-AI/PrismerCloud --skill prismer-songsee --agent claude-code

How it fires

How this skill gets triggered: by you, by Claude, or both.

  • Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
  • You can call itInvoke it directly when you want it.
  • Slash command/prismer-songsee

Context preview

The summary Claude sees to decide when to auto-load this skill.

Audio spectrograms/features (mel, chroma, MFCC) via CLI.

SKILL.md

prismer-songsee.SKILL.md
name: prismer-songsee
scope: common
category: media
description: "Audio spectrograms/features (mel, chroma, MFCC) via CLI."
version: 1.0.0
author: community
license: MIT
platforms: [ linux, macos, windows ]
metadata:
  nativeReplaces: [ songsee ]
  hermes:
    tags: [ Audio, Visualization, Spectrogram, Music, Analysis ]
    homepage: https://github.com/steipete/songsee
  requiresExplicitGrant: true
prerequisites:
  commands: [ songsee ]

songsee

Generate spectrograms and multi-panel audio feature visualizations from audio files.

Prerequisites

Requires [Go](https://go.dev/doc/install):

go install "github.com/steipete/songsee/cmd/songsee@${SONGSEE_VERSION:?Set an exact reviewed release or commit}"

Optional: `ffmpeg` for formats beyond WAV/MP3.

Quick Start

# Basic spectrogram
songsee track.mp3

# Save to specific file
songsee track.mp3 -o spectrogram.png

# Multi-panel visualization grid
songsee track.mp3 --viz spectrogram,mel,chroma,hpss,selfsim,loudness,tempogram,mfcc,flux

# Time slice (start at 12.5s, 8s duration)
songsee track.mp3 --start 12.5 --duration 8 -o slice.jpg

# From stdin
cat track.mp3 | songsee - --format png -o out.png

Visualization Types

Use `--viz` with comma-separated values:

| Type | Description | |------|-------------| | `spectrogram` | Standard frequency spectrogram | | `mel` | Mel-scaled spectrogram | | `chroma` | Pitch class distribution | | `hpss` | Harmonic/percussive separation | | `selfsim` | Self-similarity matrix | | `loudness` | Loudness over time | | `tempogram` | Tempo estimation | | `mfcc` | Mel-frequency cepstral coefficients | | `flux` | Spectral flux (onset detection) |

Multiple `--viz` types render as a grid in a single image.

Common Flags

| Flag | Description | |------|-------------| | `--viz` | Visualization types (comma-separated) | | `--style` | Color palette: `classic`, `magma`, `inferno`, `viridis`, `gray` | | `--width` / `--height` | Output image dimensions | | `--window` / `--hop` | FFT window and hop size | | `--min-freq` / `--max-freq` | Frequency range filter | | `--start` / `--duration` | Time slice of the audio | | `--format` | Output format: `jpg` or `png` | | `-o` | Output file path |

Notes

  • WAV and MP3 are decoded natively; other formats require `ffmpeg`
  • Output images can be inspected with `vision_analyze` for automated audio analysis
  • Useful for comparing audio outputs, debugging synthesis, or documenting audio processing pipelines

Execution limits

Record `songsee --help` and the installed version/build provenance. Use a task-owned output path; never overwrite input audio. Probe input duration/channels before decoding, cap file size and rendering time to the task budget, and test the actual requested codec. Core analysis is local and needs no account. A passing WAV smoke does not establish support for every codec or OS.

Read more
Ships withprismercloud

Prismer Cloud

Get the whole plugin
Stats
1,554
Stars
17
Forks
Active
Maintenance
TypeScript
Language
MIT
License
2d ago
Last commit
6mo ago
Created

Repo: Prismer-AI/PrismerCloud

Other skills on prismercloud.