An open-source, end-to-end neural voice conversion & model training ecosystem built for ultra-low latency, quality, and cross-platform flexibility.
┌───────────────────────────┐ ┌───────────────────────────┐ ┌───────────────────────────┐
│ 1. Dataset Web UI │ ───► │ 2. Model Trainer │ ───► │ 3. Real-Time Changer │
│ • Audio Slicing & Clean │ │ • macOS (MPS) & Win (CUDA│ │ • Ultra-low latency VST │
│ • Quality Validation │ │ • Free Google Colab │ │ • Live Discord/OBS stream│
└───────────────────────────┘ └───────────────────────────┘ └───────────────────────────┘
Step 1: Dataset Preparation (Web UI)
- Launch Beatrice-dataset-webui-macos or Beatrice-dataset-webui-windows depending on your OS.
- Load your raw voice recordings into the interface.
- Automatically slice long files into clean audio segments, validate transcriptions/silence, and export a ready-to-train dataset zip.
Step 2: Model Training
- macOS (Local): Use Beatrice-trainer-macos for native Apple Silicon MPS training.
- Windows (Local): Use Beatrice-trainer-windows for CUDA / DirectML hardware acceleration.
- Cloud (Free GPU): Launch Beatrice-colab on Google Colab or Kaggle to train without needing a high-end local GPU.
Step 3: Real-Time Live Inference
- Export your trained voice model.
- Open Beatrice-voicechanger-macos or Beatrice-voicechanger-windows.
- Route your physical microphone into the voice changer and output to a virtual audio channel (e.g. VB-Cable / BlackHole).
- Use transformed voices live in Discord, OBS, games, or live streaming apps!
