Home Assistant app for on-device text-to-speech. A catalog of models — Hojo TTS Light, MOSS-TTS-Nano and OmniVoice — runs on your own CPU or GPU, with built-in voices, voices designed from a set of attributes, or one cloned from a short recording. No cloud, no per-character bill. The models ship no text front-end, so the app carries one: numbers, units, clocks and dates are written out, and Chinese is converted from Traditional glyphs and respelled for Taiwan readings before synthesis. The admin UI shows that prepared text beside the composer, with a ledger of every rewrite.
Three catalog entries, one per family, none baked into the image; each is downloaded from the app's own UI. The line-up, what each costs, how each clones and which to pick is docs/models.md.
Install and start the app, then follow the App Store page, cortex-tts/DOCS.md: download a model, add the companion integration, pair, and assign a voice to a pipeline.
| Page | What it covers |
|---|---|
| App Store page | Install, configure, troubleshoot |
| Models | The line-up, what each costs, how each clones, which to pick |
| The text pipeline | Why Traditional Chinese and numbers are rewritten, and into what |
| Cloned voices | The recording, the transcript, the name |
| Keeping up | How a reply is paced, where the RTF threshold is, and the sensors |
| Running it elsewhere | A faster CPU or a GPU outside Home Assistant OS |
| HTTP API | Using the app without the integration |
| Integration | tts.speak, voice ids, speaking mode, the sensors |
Hojo TTS Light, MOSS-TTS-Nano and OmniVoice — the models this app serves. OmniVoice's ONNX graph is rhasspy/omnivoice-onnx.
MIT — see LICENSE.md.
