NVIDIA NIM Embeddings API
OpenAI-compatible embeddings endpoint (/v1/embeddings) backed by NVIDIA NeMo Retriever text embedding models including NV-Embed, NV-EmbedQA-E5, llama-3.2-nv-embedqa-1b, and BAAI BGE-M3. Returns dense float vectors for documents or queries to power RAG, semantic search, and clustering. Supports `input_type=passage|query` for asymmetric retrieval and the standard `dimensions` parameter on models that permit dimension reduction.
NVIDIA NIM Embeddings API is one of 11 APIs that NVIDIA NIM publishes on the APIs.io network, described by a machine-readable OpenAPI specification.
This API exposes 1 JSON Schema definition.
Tagged areas include AI, Artificial Intelligence, Embeddings, Retrieval, and RAG. The published artifact set on APIs.io includes API documentation, an OpenAPI specification, and 1 JSON Schema.
This API exposes 1 operation across 1 path, and defines 2 schemas. It is described by OpenAPI 3.1.0, at version 2026-05-25.
Requests are made against 2 base URLs: https://integrate.api.nvidia.com, http://localhost:8000.
Metadata
The identity and technical contract details declared by the specification.
Authentication & Security 1
NVIDIA NIM Embeddings API declares
1 security scheme
for authenticating requests.
It accepts HTTP bearer tokens (nvapi-...) (BearerAuth).
By default, every request must be authenticated.
Paths & Operations 1
Across 1 path, the API surfaces 1 operation — 1 POST. Each is listed below with its method, path, parameters, and response codes.
Dense vector embedding operations for RAG and semantic search
Schemas 2
The contract defines 2 schemas that model the data the API accepts and returns. The most detailed are EmbeddingRequest (7 properties), EmbeddingResponse (4 properties). Each schema is shown below with its type and property counts.
Specification
The full machine-readable OpenAPI contract behind this narrative.
Source
More from NVIDIA NIM 10
Other APIs NVIDIA NIM publishes across the network.