Kensho Extract

Document intelligence API that transforms PDF documents into structured JSON, with hierarchical extraction, optical character recognition (OCR), enhanced table extraction, and bounding-box locations. Supports both direct upload and a presigned upload/download URL flow for large documents.

API entry from apis.yml

apis.yml Raw ↑
name: Kensho Extract
description: Document intelligence API that transforms PDF documents into structured JSON, with hierarchical
  extraction, optical character recognition (OCR), enhanced table extraction, and bounding-box locations.
  Supports both direct upload and a presigned upload/download URL flow for large documents.
humanURL: https://docs.kensho.com/extract/home
baseURL: https://extract.kensho.com
tags:
- Document Extraction
- PDF
- OCR
properties:
- type: Documentation
  url: https://docs.kensho.com/extract/home
- type: APIReference
  url: https://docs.kensho.com/extract/api