Project overview
Run one command to turn a mixed track into instrumental, full-vocal, harmony, lead-vocal, extracted-reverb, dereverbed, and lightly de-essed outputs. The cross-platform Python driver uses several audio-separator model ensembles and can also be invoked as an agent skill. It requires Python, ffmpeg, and audio-separator, downloads roughly 5–6 GB of models on first use, and runs much more slowly without acceleration. The downloaded separation models have their own licenses, which should be reviewed separately for commercial use.
Repository facts
- Primary language
- Python
- License
- MIT
- Repository updated
- Jun 12, 2026
- Default branch
- main
Resource types
General skill
Use cases
Media, audio and video
Platforms
Claude Code, Codex, and more
Runtime
Local