A simple and efficient web application to separate vocals from music using artificial intelligence.
3.0K
A simple and efficient web application to separate audio elements (vocals, drums, bass, other instruments) from music using artificial intelligence.
If you have Docker installed:
# Download and run
docker compose up -d
Access http://localhost:8000ā .
Done! Skip to "Usage" below.
Prerequisites:
# 1. Install FFmpeg
sudo apt-get install ffmpeg # Ubuntu/Debian
# or
brew install ffmpeg # macOS
# 2. Navigate to project folder
cd voice-separator-demucs
# 3. Install dependencies
pip install -r requirements.txt
# 4. Run
python main.py
Access http://localhost:8000ā .
ā
MP3, WAV, FLAC, M4A, AAC
š Limit: 100MB per file
ā±ļø YouTube: Maximum 10 minutes
This application uses Demucs, an AI model developed by Facebook/Meta AI specifically for music source separation. It's based on deep neural networks trained on thousands of songs.
# Using docker-compose (recommended)
docker-compose up -d
# Or using Docker directly
docker build -t voice-separator .
docker run -p 8000:8000 -v $(pwd)/static/output:/app/static/output voice-separator
# With persistent model cache
docker-compose -f docker-compose.yml up -d
# Models are cached in a Docker volume for better performance
"FFmpeg not found"
# Ubuntu/Debian
sudo apt-get install ffmpeg
# macOS
brew install ffmpeg
# Windows
# Download from https://ffmpeg.org/download.html
Very slow processing
YouTube download error
Out of memory errors
# Clone repository
git clone https://github.com/paladini/voice-separator-demucs.git
cd voice-separator-demucs
# Install dependencies
pip install -r requirements.txt
# Run development server
python main.py
http://localhost:8000/docshttp://localhost:8000/redocThis tool is intended for personal and educational use. Please respect the copyright of the music you process.
Fernando Paladini (@paladiniā )
Based on the Demucs model by Facebook/Meta AI Research.
This project is licensed under the MIT License. See the LICENSE file for details.
Content type
Image
Digest
sha256:a59170d3eā¦
Size
544.2 MB
Last updated
5 months ago
docker pull paladini/voice-separator