Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text to Speech Model

TTS HD

TTS HD is a high-definition text-to-speech model from OpenAI that converts text into natural-sounding audio.

PublisherOpenAI
TypeText to Speech
Context Window4,096 tokens
ReleasedNovember 2023
Price$30.00 / 1M characters

High-quality neural text to speech audio

TTS HD (model ID: tts-1-hd) is a text-to-speech model developed by OpenAI and released in November 2023. It converts written text into spoken audio and is designed to produce higher audio quality than the standard TTS-1 model, making it suited for use cases where audio fidelity is a priority. It supports a context window of 4,096 tokens and is priced at $30.00 per one million characters.

The model offers a selection of preset voices and is intended for applications such as voiceovers, accessibility tools, content narration, and any workflow requiring natural-sounding synthesized speech. Because it is a first-party OpenAI model, it is accessible directly through the OpenAI API without requiring third-party integrations. On MindStudio, users can select a voice and pass text input to generate audio output without managing API keys directly.

What TTS HD supports

Speech Synthesis

Converts written text into spoken audio output. Supports up to 4,096 tokens of input text per request.

Voice Selection

Allows users to choose from a set of preset voices via a select input. OpenAI offers multiple distinct voice options for TTS HD.

High-Fidelity Audio

Produces audio at a higher quality level than the standard TTS-1 model. Optimized for use cases where audio clarity and naturalness are important.

Long-Form Narration

Handles extended text passages within the 4,096-token context window. Suitable for narrating articles, documents, or multi-paragraph content.

Ready to build with TTS HD?

Get Started Free

Common questions about TTS HD

What is the context window for TTS HD?

TTS HD supports a context window of 4,096 tokens, which determines the maximum amount of text that can be submitted in a single request.

How is TTS HD priced?

TTS HD is priced at $30.00 per one million characters of input text processed.

What is the difference between TTS-1 and TTS-1-HD?

TTS-1-HD is designed to produce higher audio quality than the standard TTS-1 model. TTS-1 is optimized for lower latency, while TTS-1-HD prioritizes audio fidelity.

What voices are available with TTS HD?

TTS HD supports a selection of preset voices that can be chosen via the voice select input. OpenAI offers multiple named voice options including Alloy, Echo, Fable, Onyx, Nova, and Shimmer.

When was TTS HD released?

TTS HD was released in November 2023 by OpenAI.

Parameters & options

VoiceSelect

Voice to use in TTS

Default: alloy
AlloyEchoFableOnyxNovaShimmer

Start building with TTS HD

No API keys required. Create AI-powered workflows with TTS HD in minutes — free.