How AI is applied across API Evangelist and APIs.io. Read my AI disclosure →
API Evangelist API Evangelist
Discovery
Learnings
Guidance
Toolbox
Alignment
API Evangelist LLC

Scalable Inference Serving Inference API

Model inference request endpoints

Scalable Inference Serving Inference API is one of 9 APIs that Scalable Inference Serving publishes on the APIs.io network, described by a machine-readable OpenAPI specification.

Tagged areas include Inference. The published artifact set on APIs.io includes an OpenAPI specification, API documentation, a changelog, and a getting-started guide.

This API exposes 2 operations across 2 paths, and defines 7 schemas. It is described by OpenAPI 3.1.0, at version v2.

Requests are made against a single base URL, https://inference.kserve.example.com.

2 operations 2 paths 7 schemas 2 POST

Metadata

The identity and technical contract details declared by the specification.

Specification
OpenAPI 3.1.0
API Version
v2
Server
https://inference.kserve.example.com
License
Resource Areas
1

Paths & Operations 2

Across 2 paths, the API surfaces 2 operations — 2 POST. Each is listed below with its method, path, parameters, and response codes.

Inference 2

Model inference request endpoints

POST
/v2/models/{model_name}/infer
Run Model Inference
RunInference 1 param body → 200400404503
POST
/v2/models/{model_name}/versions/{model_version}/infer
Run Model Version Inference
RunModelVersionInference 2 params body → 200400404

Schemas 7

The contract defines 7 schemas that model the data the API accepts and returns. The most detailed are InferenceResponse (5 properties), RequestInput (5 properties), ResponseOutput (5 properties), InferenceRequest (4 properties). Each schema is shown below with its type and property counts.

InferenceRequest
object
Request body for submitting model inference. Contains input tensors and optionally specifies which outputs to return.
4 properties 1 required
ResponseOutput
object
A single output tensor in the inference response.
5 properties 4 required
RequestOutput
object
Specifies which output tensor to include in the response.
2 properties 1 required
RequestInput
object
A single input tensor for an inference request.
5 properties 4 required
ErrorResponse
object
Error response returned when an inference or metadata request fails.
1 property 1 required
TensorDatatype
string
Data type of a tensor. Follows the Open Inference Protocol datatype naming convention.
InferenceResponse
object
Response from a successful model inference request.
5 properties 2 required

Specification

The full machine-readable OpenAPI contract behind this narrative.

Source

scalable-inference-serving-inference-api-openapi.yml Raw ↑

Other APIs Scalable Inference Serving publishes across the network.

BentoML REST API
vLLM OpenAI-Compatible API
NVIDIA Triton Inference Server HTTP API
MLflow Model Registry REST API
Ray Serve REST API
Scalable Inference Serving Health API
Scalable Inference Serving Metadata API
Scalable Inference Serving Models API
Where this information came from

This is an independent, third-party profile of Scalable Inference Serving Inference API, published by API Evangelist. We do not operate, host, resell, or support these APIs, and we are not affiliated with or endorsed by the company unless stated above. Everything here is built from publicly available information — the company's own site, developer portal, documentation, public repositories, and the specifications it publishes for public use. Nothing is obtained by breaching a system, defeating an access control, or using credentials.

The Kin Score and Agent Readiness rating are independently calculated assessments of a company's public API artifacts, scored against a published rubric. They are not certifications, endorsements, security assessments, or audits.

Corrections, re-scores, and removal are free — no partnership or purchase required, and you do not need to justify the request. A removed company is recorded as unrated, never scored zero for having asked. Acknowledgement within one business day; removal within two.

info@apievangelist.com · Read the full data-sourcing policy →
On a security or compliance team? Put security in the subject line and you will get a person, not a form — we will tell you exactly which public URLs this profile was built from.