AI Voice Glossary
Plain-English definitions for TTS, STT, voice cloning, voice design, points, SRT, temperature and MCP — including why there's no pitch control.
Step-by-step guide
How to Use AI Voice Glossary
Start for free — no credit card, no software to install.
Start with TTS and STT
TTS converts text into speech; STT converts speech into a plain transcript — opposite directions, and STT here returns no captions or speaker labels.
Understand voice cloning vs voice design
Cloning reuses a real sample of a specific voice with consent; design builds a new voice from a written description, with preview takes before saving it.
Learn what a point buys
A point pays for about one second of generated audio, at a slightly different rate in the app than through the public API.
Know the two sliders
Speed changes pacing; Temperature changes delivery variation. There is no pitch control — voice character comes from voice choice, not a slider.
Why creators choose this tool
AI Voice Glossary — Features & Capabilities
Everything you need to create professional voiceovers — in one tool.
TTS and STT defined
Text-to-speech generates audio from text; speech-to-text returns a plain transcript from audio, with no captions or timings.
Cloning vs design defined
Cloning uses a real voice sample with consent; design builds a new voice from a plain-English description.
Points and SRT defined
A point pays for roughly one second of audio; SRT is a caption file built from the word-level timing the speech engine returns.
Temperature and MCP defined
Temperature controls delivery variation, not pitch — there is no pitch slider. MCP lets an AI assistant use the platform's voice tools directly.
Points in the app versus the API
The app spends points at roughly one per 17 characters of text; the public API uses a slightly different, slightly higher rate per character.
Use cases
What Can You Create with AI Voice Glossary?
From TikTok to podcasts — here's where creators use this tool most.
Understanding a confusing term before asking support
Looking up what SRT or MCP actually means before opening a help ticket.
Explaining the tool to a non-technical teammate
Sharing plain definitions instead of jargon when introducing the platform to someone new.
Deciding which feature actually solves a problem
Realizing 'I want a higher voice' means picking a different voice, not adjusting a pitch slider that doesn't exist.
Evaluating whether MCP access is relevant
Understanding what MCP does before deciding whether an assistant-integrated workflow is worth the Pro plan.
vs Standard TTS
Vocallab vs Standard Text-to-Speech
See why creators switch from generic TTS tools to Vocallab.
About this tool
Everything you need to know about AI Voice Glossary
The terms in this space get used loosely, so here's what each one actually means on VocalLab. TTS, text-to-speech, turns written text into spoken audio; STT, speech-to-text, does the reverse, returning a plain transcript with no timings or speaker labels. A point is the unit that pays for generated audio, spent at roughly one per second. SRT is a caption file built from word-level timing. Temperature controls how varied a delivery sounds, not pitch — there's no pitch control at all, only Speed and Temperature. MCP is a way for an AI assistant to use the platform's voice tools directly instead of a person clicking through the app.
People also search for
FAQ
AI Voice Glossary — FAQ
Common questions users ask about the AI Voice Glossary voice. If you need additional help, please contact us via the contact form or email us at support@vocallab.ai.
Is there a pitch control?
No — the two available sliders are Speed and Temperature; pitch isn't adjustable on any voice, so a higher or lower sound comes from picking a different voice.
What does Temperature actually control?
How varied the delivery sounds from one generation to the next — it's not a tone or pitch setting, despite the name sounding like one.
What's the difference between voice cloning and voice design?
Cloning reuses a real person's voice from a short sample with consent; design builds an entirely new voice from a written description, with preview takes before you save one.
What is MCP in this context?
A hosted server that lets an AI assistant call the platform's voice tools directly — listing voices, using a clone, or generating speech — instead of a person doing it through the app.
Related AI Voice Tools
Turn Text Into Humanlike Speech
Generate humanlike speech for voice assistants, IVR prompts, and conversational apps. Download MP3, adjust tone, free to try with no signup.
GeneralGenerate Voiceovers for Content Creators
Generate one script into voiceovers sized for YouTube, TikTok, and podcasts — pick tone per platform, download MP3 and SRT, no editing suite needed.
GeneralTurn Content Into Voiceover Audio
Layer a narrated voiceover track over footage you've already filmed. Paste your narration script, download MP3 timed to drop into your video edit.
GeneralGenerate Natural Sounding Voice from Text
Generate a natural sounding voice for long-form narration like audiobooks and full articles. Download MP3 and SRT, free to try, no signup.
GeneralGenerate Content Using Your Voice
Pick a single go-to AI voice and reuse it across every video, podcast, and post you publish, so your content always sounds like you. Free to try.
GeneralCreate AI Voice Avatar
Give a virtual presenter, digital assistant, or streaming persona a consistent AI voice. Pick a character voice, download MP3, free to try.


