# Quick Start Guide
Get up and running with Karaoke Maker in 5 minutes!
# Prerequisites
- Python 3.8+: Check with
python3 --version - FFmpeg: Required for video encoding
- Linux:
sudo apt install ffmpeg(Ubuntu/Debian) orsudo dnf install ffmpeg(Fedora) - macOS:
brew install ffmpeg - Windows: Download from ffmpeg.org
- Linux:
# Installation
# Linux/macOS
# Navigate to project directory
cd karaoke-maker
# Run setup script (installs dependencies)
./setup.sh
# Run the application
./run.sh
# Windows
REM Navigate to project directory
cd karaoke-maker
REM Run setup script (installs dependencies)
setup.bat
REM Run the application
run.bat
# First Run
- Model Download: On first run, Whisper AI model will download (~140MB). This is one-time only.
- Wait: Initial setup may take 2-5 minutes depending on your internet speed.
- GPU: If you have an NVIDIA GPU with CUDA, it will be auto-detected for faster processing.
- Vocal Separation: By default, Demucs is not installed (requires compilation). The app will work fine without it, using the original audio for transcription.
# Optional: High-Quality Vocal Separation
For best results, install Demucs after initial setup:
Linux:
sudo apt install lame liblame-dev # Ubuntu/Debian
# or
sudo dnf install lame lame-devel # Fedora
# Then activate venv and install
source venv/bin/activate
pip install lameenc
pip install -U git+https://github.com/facebookresearch/demucs#egg=demucs
macOS:
brew install lame
source venv/bin/activate
pip install lameenc
pip install -U git+https://github.com/facebookresearch/demucs#egg=demucs
# Basic Usage
# Method 1: Auto-Transcribe (Recommended)
- Click Browse and select your MP3 file
- Keep Auto-transcribe selected
- Choose a background gradient
- Click Generate Karaoke Video
- Choose save location
- Wait 3-5 minutes for processing
- Done!
# Method 2: Online Lookup
- Click Browse and select your MP3 file
- Select Lookup online
- Artist and title should auto-fill from MP3 metadata
- Click Lookup Lyrics
- Review lyrics in the text area
- Choose a background gradient
- Click Generate Karaoke Video
- Wait 2-4 minutes for processing
- Done!
# Method 3: Manual Entry
- Click Browse and select your MP3 file
- Select Enter manually
- Paste or type your lyrics in the text area
- Choose a background gradient
- Click Generate Karaoke Video
- Wait 2-4 minutes for processing
- Done!
# Tips
- Best Method: Auto-transcribe is the most reliable (works 95% of the time)
- Lyrics Lookup: AZLyrics lookup is often blocked - use Auto-transcribe or Manual instead
- Better Accuracy: Auto-transcribe works best with clear vocals
- Fast Processing: Use GPU if available (10-15x faster)
- Custom Backgrounds: Click “Custom Image” to upload your own background
- Shorter Songs: Test with shorter songs (2-3 minutes) first
- Check Output: Generated videos are saved as MP4 files
# Troubleshooting
# “FFmpeg not found”
- Install FFmpeg and ensure it’s in your system PATH
- Test with:
ffmpeg -version
# Out of memory
- Close other applications
- Try shorter songs
- Use CPU-only mode (disable CUDA)
# Lyrics not found
- Check artist/title spelling
- Try auto-transcribe instead
- Use manual entry as fallback
# Slow processing
- First run downloads models (one-time)
- Older GPUs may not work - app auto-switches to CPU (see GPU_TROUBLESHOOTING.md)
- CPU mode: 6-8 minutes for 4-minute song
- Compatible GPU: 1.5-2 minutes for 4-minute song
- Shorter songs process faster
# Next Steps
- Read the full README.md for detailed information
- Check CONTRIBUTING.md if you want to contribute
- Report issues on GitHub
# Need Help?
- Check the README.md for detailed documentation
- Open an issue on GitHub
- Review the CONTRIBUTING.md for development info
Enjoy creating karaoke videos!