Time-accurate speech recognition with word-level timestamps, speaker diarization, and multi-speaker transcription powered by Whisper.
# 1. Clone repository git clone https://github.com/m-bain/whisperX.git cd whisperx # 2. Set up virtual environment python3 -m venv .venv source .venv/bin/activate # 3. Install dependencies & configure env pip install -r requirements.txt cp .env.example .env # 4. Start local development server python main.py
The original and most feature-complete open-source browser interface for Stable Diffusion and SDXL generation with support for ControlNet and extensions.