Properties
| Name | Type | Description |
|---|---|---|
| model | string | One of the available TTS models. tts-1 is optimized for speed, tts-1-hd is optimized for quality, and gpt-4o-mini-tts supports advanced voice instructions. |
| input | string | The text to generate audio for. The maximum length is 4096 characters. |
| voice | string | The voice to use when generating the audio. Previews of the voices are available in the Text to Speech guide. |
| instructions | string | Control the voice of your generated audio with additional instructions. Only supported with gpt-4o-mini-tts. |
| response_format | string | The format to audio in. Supported formats are mp3, opus, aac, flac, wav, and pcm. Opus is recommended for internet streaming and communication, aac for digital audio compression, and flac for lossless |
| speed | number | The speed of the generated audio. Select a value from 0.25 to 4.0. 1.0 is the default. |
JSON Schema
{
"$schema": "https://json-schema.org/draft/2020-12/schema",
"title": "CreateSpeechRequest",
"type": "object",
"properties": {
"model": {
"type": "string",
"description": "One of the available TTS models. tts-1 is optimized for speed, tts-1-hd is optimized for quality, and gpt-4o-mini-tts supports advanced voice instructions."
},
"input": {
"type": "string",
"description": "The text to generate audio for. The maximum length is 4096 characters."
},
"voice": {
"type": "string",
"description": "The voice to use when generating the audio. Previews of the voices are available in the Text to Speech guide."
},
"instructions": {
"type": "string",
"description": "Control the voice of your generated audio with additional instructions. Only supported with gpt-4o-mini-tts."
},
"response_format": {
"type": "string",
"description": "The format to audio in. Supported formats are mp3, opus, aac, flac, wav, and pcm. Opus is recommended for internet streaming and communication, aac for digital audio compression, and flac for lossless audio compression."
},
"speed": {
"type": "number",
"description": "The speed of the generated audio. Select a value from 0.25 to 4.0. 1.0 is the default."
}
}
}
Work with this as data
Every JSON Schema here is available over the APIs.io API and to AI agents over MCP.
MCP server
One button, every client — Claude, Cursor, VS Code and the rest.
https://apis.io/mcp
Tools for schemas
4 MCP tools reach this
find_json_schemasBrowse and filter every JSON Schema in the catalog.apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.resolveTurn a domain, URL or GitHub org into the provider it belongs to.find_cohortsEvery scored population of providers in the catalog.
Call it yourself
curl for this page
This JSON Schema
curl "https://apis.io/api/v1/json-schemas/openai-audio-create-speech-request"
All schemas
curl "https://apis.io/api/v1/json-schemas?limit=25"
Discovery needs no key. Ratings and market analysis are Pro.
Get an API key
Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.
A second provider on the same verified email joins the account you already have.