Local models
| Component | Tested model/version | Approximate model assets |
|---|---|---|
| English speech | Supertonic 3, F4; Python supertonic==1.3.1 | 381 MiB |
| Speech recognition | SenseVoice, sherpa-onnx int8, 2024-07-17 | 228 MiB |
| Language model | Ollama qwen3:1.7b | 1.3 GB |
| Desktop shell | Electron 44.7.0, arm64 | Separate app download |
Model weights are downloaded during setup, rather than included in Git. Supertonic’s tested revision is pinned in scripts/download_models.py; the SenseVoice ONNX download is verified by SHA-256. Ollama’s model tag may change upstream.
To rebuild the demos, record both clips in Settings → Demo recording, install ffmpeg (brew install ffmpeg), and run ./scripts/export_media.command. The included introduction audio can be regenerated with .venv/bin/python scripts/generate_intro.py.
Source captured: 2026-10-11