v1.0 — initial release
- First public version of the site: text-to-speech and voice cloning available.
A plain-language guide to using this voice studio — no technical background needed.
← Back to the appVoiceStudio turns written text into spoken audio, and can also learn the sound of a specific voice from a short recording so it can speak new sentences in that voice. Everything runs on this private server — nothing you type or upload is sent to a third party unless you specifically choose a "Cloud" engine.
This teaches the app to speak in a specific person's voice, using a short sample recording as a reference.
This server uses a single access key to protect administrative functions. Treat it like a password: don't share it, paste it into chat messages, or post it anywhere public. If you're just asked to try the site casually, use the share link/PIN you were given instead of the main key.
This site is currently reachable from the public internet. That's a deliberate, temporary setup while we're getting it running — it will later be restricted to trusted devices only.
This server processes audio using its CPU rather than a graphics card, so generation — especially longer clips or voice cloning — can take noticeably longer than you may have seen elsewhere. This is expected.
Browser sessions expire after a period of inactivity for security. Just re-enter the access key you were given.
Try right-clicking the play button and choosing "Save audio as...", or reload the page and generate again.
Reload the page first. If it still doesn't work, note what you were doing and let the site owner know.