OpenAI · Schema

CreateTranscriptionRequest

Artificial IntelligenceLarge Language ModelsT1

Properties

Name Type Description
file string The audio file object to transcribe, in one of these formats: flac, mp3, mp4, mpeg, mpga, m4a, ogg, wav, or webm. File uploads are limited to 25 MB.
model string ID of the model to use. Only whisper-1 and gpt-4o-transcribe are currently available.
language string The language of the input audio. Supplying the input language in ISO-639-1 format will improve accuracy and latency.
prompt string An optional text to guide the model's style or continue a previous audio segment. The prompt should match the audio language.
response_format string The format of the transcript output. Defaults to json. verbose_json includes additional metadata like word-level timestamps.
temperature number The sampling temperature, between 0 and 1. Higher values like 0.8 will make the output more random, while lower values like 0.2 will make it more focused and deterministic.
timestamp_granularities array The timestamp granularities to populate for this transcription. response_format must be set to verbose_json to use this parameter.
View JSON Schema on GitHub

JSON Schema

openai-audio-create-transcription-request-schema.json Raw ↑
{
  "$schema": "https://json-schema.org/draft/2020-12/schema",
  "title": "CreateTranscriptionRequest",
  "type": "object",
  "properties": {
    "file": {
      "type": "string",
      "description": "The audio file object to transcribe, in one of these formats: flac, mp3, mp4, mpeg, mpga, m4a, ogg, wav, or webm. File uploads are limited to 25 MB."
    },
    "model": {
      "type": "string",
      "description": "ID of the model to use. Only whisper-1 and gpt-4o-transcribe are currently available."
    },
    "language": {
      "type": "string",
      "description": "The language of the input audio. Supplying the input language in ISO-639-1 format will improve accuracy and latency."
    },
    "prompt": {
      "type": "string",
      "description": "An optional text to guide the model's style or continue a previous audio segment. The prompt should match the audio language."
    },
    "response_format": {
      "type": "string",
      "description": "The format of the transcript output. Defaults to json. verbose_json includes additional metadata like word-level timestamps."
    },
    "temperature": {
      "type": "number",
      "description": "The sampling temperature, between 0 and 1. Higher values like 0.8 will make the output more random, while lower values like 0.2 will make it more focused and deterministic."
    },
    "timestamp_granularities": {
      "type": "array",
      "description": "The timestamp granularities to populate for this transcription. response_format must be set to verbose_json to use this parameter."
    }
  }
}

Work with this as data

Every JSON Schema here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for schemas

4 MCP tools reach this
  • find_json_schemasBrowse and filter every JSON Schema in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This JSON Schema
curl "https://apis.io/api/v1/json-schemas/openai-audio-create-transcription-request"
All schemas
curl "https://apis.io/api/v1/json-schemas?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.

A second provider on the same verified email joins the account you already have.