Kensho Extract API
Transforms unstructured PDF and image documents into machine-readable JSON, identifying titles, subtitles, paragraphs, tables, and footers in natural reading order. Optional OCR and Figure Extraction (FigEx). REST API at extract.kensho.com with asynchronous extractions and presigned upload/download URLs.
Kensho Extract API is one of 9 APIs that S&P Global publishes on the APIs.io network, described by a machine-readable OpenAPI specification.
Tagged areas include Document Extraction, OCR, PDF, Tables, and Unstructured Data. The published artifact set on APIs.io includes an OpenAPI specification, API documentation, an API reference, a getting-started guide, authentication docs, and a JSON-LD context.
This API exposes 5 operations across 5 paths, and defines 3 schemas. It is described by OpenAPI 3.0.2, at version 3.0.0.
Requests are made against a single base URL, https://extract.kensho.com/.
Metadata
The identity and technical contract details declared by the specification.
Authentication & Security 1
Kensho Extract API declares
1 security scheme
for authenticating requests.
It accepts HTTP bearer tokens (JWT) (bearerAuth).
Paths & Operations 5
Across 5 paths, the API surfaces 5 operations — 2 GET, 2 POST, 1 PUT. Each is listed below with its method, path, parameters, and response codes.
Schemas 3
The contract defines 3 schemas that model the data the API accepts and returns. The most detailed are ContentTree (4 properties), Output (2 properties). Each schema is shown below with its type and property counts.
Specification
The full machine-readable OpenAPI contract behind this narrative.
Source
More from S&P Global 8
Other APIs S&P Global publishes across the network.