Kensho Scribe
Asynchronous speech-to-text transcription API that turns audio and video into text with high accuracy. Supports batch, real-time, and human-in-the-loop transcription, multipart and remote-URL submission, HEAD polling for completion, and webhook callbacks.