S&P Global · OpenAPI Overlay 1.0.0
API Evangelist conversational phrasing for Kensho Extract Extractions API
6 actions
6 updates
phrasing
extends
openapi/sp-global-extractions-api-openapi.yml
Generated by API Evangelist
Written by API Evangelist tooling for S&P Global's API. It is a proposal applied on top of the contract, not a document S&P Global publishes.
What the actions change
x-apievangelist-phrasing
Targets 6
$.info
$.paths['/v3/extractions'].post
$.paths['/v3/extractions/upload-url'].post
$.paths['/v3/extractions/upload-complete'].put
$.paths['/v3/extractions/{request_id}'].get
$.paths['/v3/extractions/download-url/{request_id}'].get
OpenAPI Overlay
# Generated by API Evangelist (build-phrasing.py). Our phrasing, not observed demand.
overlay: 1.0.0
info:
title: API Evangelist conversational phrasing for Kensho Extract Extractions API
version: 1.0.0
extends: openapi/sp-global-extractions-api-openapi.yml
actions:
- target: $.info
update:
x-apievangelist-phrasing:
method: generated
generated: '2026-10-01'
generator: build-phrasing.py
label: Generated by API Evangelist
operations: 5
- target: $.paths['/v3/extractions'].post
update:
x-apievangelist-phrasing:
intent: Submit a document directly for structured extraction
effect: write
questions:
- How do I send a PDF straight to Kensho Extract and get its titles, paragraphs and tables back?
- Can I extract only certain pages of a document, like pages 1-5 and 7?
- What is the largest file I can attach directly in a single extraction request?
instructions:
- text: Attach {file} directly and extract it as a {document_type} document with OCR set to {ocr} and enhanced tables {enhanced_table_extraction}.
slots:
file: requestBody.file
document_type: requestBody.document_type
ocr: requestBody.ocr
enhanced_table_extraction: requestBody.enhanced_table_extraction
- text: Upload {file} in the request body and extract only pages {pages}, also returning figures ({figure_extraction}).
slots:
file: requestBody.file
pages: requestBody.pages
figure_extraction: requestBody.figure_extraction
- text: Send {file} for extraction tagged with my own document ID {document_id} at {priority} priority, including image locations ({include_images}).
slots:
file: requestBody.file
document_id: requestBody.document_id
priority: requestBody.priority
include_images: requestBody.include_images
method: generated
generated: '2026-10-01'
- target: $.paths['/v3/extractions/upload-url'].post
update:
x-apievangelist-phrasing:
intent: Get a pre-signed URL to upload a document for extraction
effect: write
questions:
- Can I get a pre-signed upload link instead of attaching my document to the extraction request?
- Which settings do I have to choose before I receive a pre-signed URL for my document?
- Is there a way to cap the total number of pages extracted from a document I upload by link?
instructions:
- text: Get a pre-signed upload URL for a {document_type} document as {output_format}, OCR {ocr}, enhanced tables {enhanced_table_extraction}.
slots:
document_type: requestBody.document_type
output_format: requestBody.output_format
ocr: requestBody.ocr
enhanced_table_extraction: requestBody.enhanced_table_extraction
- text: Get an upload link for a document I'll send separately, extracting at most {num_pages_to_extract} pages.
slots:
num_pages_to_extract: requestBody.num_pages_to_extract
- text: Reserve a pre-signed upload URL for document {document_id} at {priority} priority.
slots:
document_id: requestBody.document_id
priority: requestBody.priority
method: generated
generated: '2026-10-01'
- target: $.paths['/v3/extractions/upload-complete'].put
update:
x-apievangelist-phrasing:
intent: Start extraction after uploading to the pre-signed URL
effect: write
questions:
- Why hasn't extraction started even though my file finished uploading to the pre-signed link?
- What do I call to tell the service my presigned upload is done?
instructions:
- text: Mark the upload for extraction request {request_id} as complete so processing begins.
slots:
request_id: requestBody.request_id
- text: I've finished uploading to the pre-signed URL for {request_id}; kick off the extraction.
slots:
request_id: requestBody.request_id
method: generated
generated: '2026-10-01'
- target: $.paths['/v3/extractions/{request_id}'].get
update:
x-apievangelist-phrasing:
intent: Retrieve the extracted document for a request
effect: read
questions:
- Where do I fetch the structured output once my document has been extracted?
- Can I get the extracted content back with character offsets or element locations included?
instructions:
- text: Get the extracted document content for request {request_id}.
slots:
request_id: path.request_id
- text: Return the extraction result for {request_id} inline as {output_format}.
slots:
request_id: path.request_id
output_format: query.output_format
method: generated
generated: '2026-10-01'
- target: $.paths['/v3/extractions/download-url/{request_id}'].get
update:
x-apievangelist-phrasing:
intent: Get a download link for an extracted document
effect: read
questions:
- Is there a file download link for a finished extraction rather than the inline response?
- Which field holds the URL I use to download my extracted output?
instructions:
- text: Give me the output_url to download the extracted document for request {request_id}.
slots:
request_id: path.request_id
- text: Fetch the download link for the finished extraction {request_id}.
slots:
request_id: path.request_id
method: generated
generated: '2026-10-01'