On-device speech synthesis. No audio sent to a server. Also the listening-test tool for picking the THOX default voice.
local-firstprobing…Open RAIL-Mno default picked
1 · Load
WebGPU
…
Model
onnx-community/Supertonic-TTS-ONNX · fp32 only
First load
262.9 MB (one time, then cached)
Output
44.1 kHz mono · English only
Not loaded. Nothing is downloaded until you press Load.
2 · Text
–RTF ↓
–synth ms
–audio s
–backend
3 · Listening test — pick the THOX default
All 10 shipped voices. Click one to hear it, then Set default.
This exists because no naturalness claim may be made until a human listens —
open item #9 in the review.
Licence obligations (BigScience Open RAIL-M).
This audio is machine-generated and any THOX surface presenting it must say so
(Attachment A(e)). The use restrictions must be passed to end users (§5) and are
viral — a rebrand cannot strip them. No voice cloning: only these 10 stock
voices ship, and no speaker-embedding upload is exposed (Attachment A(g)).