en_GB-vctk-medium
en_GB-vctk-medium is a English (English) voice for the Piper
engine, and Quick TTS runs it free in your browser with nothing uploaded. This page is
the honest version: what tier it is, what sits beneath it when your device cannot load
it, what the alternative on English is, and where it falls short.
What this voice actually is
| Fact | Value |
|---|---|
| Voice ID | en_GB-vctk-medium |
| Engine | Piper, running as WebAssembly in your browser |
| Language | English (English) |
| Quality tier | medium, the tier Quick TTS treats as the default everywhere, and the best quality-per-megabyte Piper offers |
| Preselected? | No. en_US-libritts_r-medium is the English default; pick this one from the voice list. |
| Needs a GPU? | No. Piper is the engine that runs without one. |
| Step-down tier | None. This voice is single-rung, so a device that cannot load it falls back to the browser engine rather than stepping down to a smaller model. |
| Also available in HD? | Yes. Supertonic HD speaks English at 44.1kHz, on WebGPU only. |
| Integrity | SHA-256 pinned. Weights are rejected unless their hash matches the pin, on every host we fetch from. |
What is worth knowing about it
VCTK is a multi-speaker corpus, which makes this the broadest-sounding of the three English
voices, but Quick TTS exposes it as one selectable voice, not as a speaker picker. If you
want a British English read rather than an American one, this is the option. If you want the
most natural American English long-form read, en_US-libritts_r-medium is the
default for a reason.
The other English Piper voices
English ships 3 selectable Piper voices, so this one is a choice rather than the only option:
en_US-libritts_r-medium(the preselected default)en_US-joe-medium
What happens when your device cannot load it
Piper's real failure mode is memory, not speed. The model has to allocate an ONNX session, and on weak or GPU-less hardware that allocation is the thing that fails. Quick TTS handles it in two layers, and both are automatic:
- Before the download. A device reporting 4GB of memory or less starts on the smallest available tier instead of the default, so it never pulls the large model only to fail on it. Below 2GB, Piper is hidden entirely rather than offered and then failing.
- After a failure. English has no smaller working model, so there is nothing to step down to. The remainder of the text is handed back to the browser's built-in speech engine instead, and Supertonic HD is the other neural option if the machine has a working GPU.
How to use it
Open quick-tts.com,
switch the engine dropdown to Piper, and choose en_GB-vctk-medium from the voice
list. The neural voice list follows the page language, so English voices appear on
the English page.
Speed, volume and voice can all be changed mid-read and the current passage
replays under the new settings without starting over. Switching the engine
itself stops playback, so you press play once to start the new one. There is no
mid-read handover of the remaining text between engines.
The first play click downloads the model once and caches it; afterwards it starts immediately and keeps working offline.
Nothing you paste is uploaded
Piper runs as WebAssembly inside your own browser tab. The model is downloaded once and cached; after that, the text you paste is turned into audio on your machine and never sent anywhere. That is an architectural property rather than a policy promise: the test is that generation keeps working with your network disconnected once the model has loaded.
Common questions
What is en_GB-vctk-medium?
It is a English neural text-to-speech voice for the Piper engine, at the medium quality tier. Quick TTS runs it as WebAssembly inside your browser, so it needs no GPU and no account, and the text you paste is never uploaded.
Is en_GB-vctk-medium free?
Yes. There is no account, no character limit and no watermark. The model downloads once to your browser and then runs on your own device, so there is no per-character server cost to meter.
Do I need a GPU to use en_GB-vctk-medium?
No. Piper runs on WebAssembly and works on machines with no usable GPU at all, which is why it is the most-selected engine here. The GPU-only alternative on English is Supertonic HD.
How do I select en_GB-vctk-medium in Quick TTS?
Open quick-tts.com, switch the engine dropdown to Piper, and pick it from the voice list. The neural voice list follows the page language, so you need the English page to see English voices.
Can I download the audio from en_GB-vctk-medium?
Yes, as WAV or MP3, generated on your device. Anything past roughly 20,000 characters is packaged as a ZIP of audio parts so that memory stays bounded on long documents.
Read next
- Every Piper voice we ship, ranked: the same voices side by side, with what each one is good at.
- All 16 selectable Piper voices
- What "in-browser TTS" actually means