Kensho Scribe

Asynchronous speech-to-text transcription API that turns audio and video into text with high accuracy. Supports batch, real-time, and human-in-the-loop transcription, multipart and remote-URL submission, HEAD polling for completion, and webhook callbacks.

API entry from apis.yml

apis.yml Raw ↑
name: Kensho Scribe
description: Asynchronous speech-to-text transcription API that turns audio and video into text with high
  accuracy. Supports batch, real-time, and human-in-the-loop transcription, multipart and remote-URL submission,
  HEAD polling for completion, and webhook callbacks.
humanURL: https://docs.kensho.com/scribe/v2/developer-guide
baseURL: https://scribe.kensho.com
tags:
- Speech to Text
- Transcription
- Audio
properties:
- type: Documentation
  url: https://docs.kensho.com/scribe/v2/developer-guide
- type: APIReference
  url: https://docs.kensho.com/scribe/v2/batch-api-specification