Skip to content
TipPage Docs
Esc
navigateopen⌘Jpreview
On this page

TTS

The voice that reads tips on stream - queue, history, voice settings, and AI voices.

The TTS page runs everything about the voice on your stream: the queue of messages waiting to play, the history of what’s played, and the voice itself.

Queue and history

The Queue panel shows what’s waiting, in order. The History panel shows what’s already played, with a replay button - great for “wait, what did that say?” moments. Replayed tips show a gold “(Replay)” badge on the overlay so chat knows it’s a rerun.

Voice settings

These settings shape the audio that gets generated for every tip:

  • TTS enabled - the master switch for generating tip audio.
  • Language - the voice, e.g. English (UK).
  • Read donor names / amounts / messages - three toggles for what gets spoken. Turn amounts off if you don’t want numbers read out.
  • Preview - plays a sample so you can tune it without firing test alerts.

Playback lives with each overlay: the per-overlay TTS toggle, volume, and the tip alert sound (including uploading and picking sound files) are all in the overlay editor, so two overlays can behave differently.

Viewers get a Preview TTS button on your tip page too, so they hear exactly how their message will sound before paying. Same voice, same settings.

AI voices

AI voices are in closed beta: they need TipPage+ and an invite from us. If the AI voices panel isn’t on your TTS page, your account isn’t in the beta yet - email hello@tippage.com if you’d like in.

When it’s on, viewers can have their tip read out in an AI voice instead of the standard one - and they can write a script, where different lines of the message are spoken by different voices, played back-to-back as one clip.

On the tip page, viewers see a “Read it in AI voices” button under the message box. It opens a script editor: each line has a voice attached, tapping the voice name opens a picker with a playable sample for every voice, and pressing Enter starts the next line for a back-and-forth. The whole thing is voiced in one go - the model performs the exchange, including the timing between turns - with up to five different voices in a message.

Delivery tags

Under the script editor there’s an Advanced row: Delivery tags. These are short directions a viewer can drop into a line to tell a voice how to say the next bit - [excited], [whisper], [sarcastic] - or to make a noise instead of a word: [laughing], [sigh], [gasp], [pause]. Tapping a chip inserts the tag where the viewer is typing.

Tags are direction, not speech. They aren’t read out, they never appear in the tip’s message on your overlay, in chat or in your history, and they don’t count towards the viewer’s character limit. There’s a cap of twelve per message.

The list of tags is fixed and we curate it - viewers can’t invent their own. Anything else typed in square brackets is treated as ordinary words and read aloud like the rest of the message, so it still goes through your word filter exactly like anything else a viewer writes.

Your controls in the panel:

  • Offer AI voices to donors - the master switch.
  • Mode - Donors can choose keeps the standard voice as the default, with AI voices as an option. All TTS uses AI voices removes the standard voice entirely: every tip message is AI-voiced.
  • Minimum tip for AI voices - set it equal to your minimum tip for no extra requirement, or higher to make AI voices a premium that viewers unlock by tipping more.
  • Voices - the voices viewers can choose from, with a play button to sample each one. Manage voices opens the two lists: This channel (voices you cloned, plus any you added from the library) and TipPage voices (ours, on by default - switch off the ones you don’t want, or all of them).

Your own voices

Manage voices on the AI voices panel opens your voice library, where you can clone a voice from a recording:

  1. Add a voice, give it a photo, a name and a short description (the name and description are what viewers see in the picker). The photo is square - pick any image and we crop and shrink it for you - and it’s how you tell your voices apart in the library, next to the name of whoever added it.
  2. Add up to 90 seconds of audio, in as many as ten clips. Clean, single-speaker speech works best: no music, no background chatter, 10 to 30 seconds at the energy you want. Shorter is also friendlier for back-and-forth messages - every voice in a message shares one budget, so a couple of 90-second voices can’t be cast together.
  3. Optionally type what’s said in each clip. Leave it blank and we’ll transcribe it for you.
  4. Tick the rights confirmation - it has to be your own voice or one you have permission to use - and hit Create voice.

Your uploads are deleted as soon as the voice is built - the voice itself is what we keep, and it lives on our own machines. Building takes a minute or two; the row says Building until it’s ready. Once it is, we record a short sample in the new voice (“Hi, this is a preview of my voice…”) and store it - that’s the clip you and your viewers hear when you tap play in the picker, so auditioning a voice is instant and never waits on the AI. Once a voice is built it’s finished - there’s nothing to tune or re-run. A failed build tells you why and offers Try again.

Voices start private to your channel. Publish lists one in the shared voice library so other streamers can find it - it doesn’t appear on their tip pages unless they choose to add it, and Unpublish takes it off the list (and out of every channel that added it). Both ask you to confirm first, since either one changes what other channels can do with the voice.

Browse voices opens that library: everything other streamers have published, each with a sample to play and the name of whoever uploaded it. Add the ones you want and they join This channel; Remove takes one back out without affecting anyone else. TipPage’s own voices aren’t in there - they’re already on, in their own list, and you switch them off rather than adding them.

On your tip page, viewers see the same split: This channel for your voices and the ones you added, Other voices for TipPage’s.

Impersonation is the one hard line: we remove cloned voices that get reported as someone who didn’t agree to it.

A few things to know:

  • Free messages can use AI voices too - but only while you haven’t set a higher minimum tip for them. A free message can’t clear a paid gate.
  • Your word filter and AI tip filter still apply. If a filter changes the message, the AI audio speaks the filtered text - or falls back to the standard voice if it can’t.
  • AI audio takes a little longer to generate than standard TTS, especially multi-voice scripts. The overlay holds the tip at the front of the queue until its audio is ready rather than playing it in the wrong voice.
  • Deleting a voice that a waiting tip used means that tip gets read in your standard voice instead.
  • Anything that stops the AI audio being made - a deleted voice, a busy GPU - falls back to your standard voice. The alert still plays.
  • There’s no message preview for viewers, but every voice has a sample clip they can play before choosing.
  • Automating it? The developer API can read the catalog and toggle the feature or individual voices.

Mod control from chat

Your mods don’t need dashboard access to run the queue - !tts skip, !tts pause, !tts resume, and !tts replay all work from chat once the bot is enabled.

Was this page helpful?