AI-Powered · Open Source · Portfolio Project
Not just 4 tracks — StemForge AI splits any song into lead vocal, backing vocals, kick, snare, hi-hat, cymbals, bass, guitar and synth in a two-stage AI pipeline.
Two-stage hierarchical pipeline — Demucs handles the main separation, then custom spectral models go deeper into drums and vocals.
Live AI processing requires ~5–15 min/song on CPU or a dedicated GPU. Not suitable for shared hosting. Backing vocal separation is experimental — mid-side masking works best on professionally mixed stereo tracks. Drum separation uses spectral methods; LarsNet integration is planned for better component isolation. Future work includes GPU support, user accounts, and a public API.