Project overview
Give it an MP3, WAV, FLAC, M4A, or OGG file to separate vocals, transcribe lyrics, and extract musical features such as BPM, key, and spectral characteristics. An AI agent then turns the raw results into genre, mood, instrument, vocal, and production analysis. Python and ffmpeg are required; a CUDA GPU is recommended but not mandatory.
Repository facts
- Primary language
- Python
- License
- MIT
- Repository updated
- May 24, 2026
- Default branch
- master
Resource types
General skill
Use cases
Media, audio and video
Platforms
Claude Code, Codex, and more
Runtime
Command line
Capabilities
Audio processingTranscription