Need help with your APIs? I offer API discovery, governance & evangelism services. Explore services →
API Evangelist API Evangelist
Discovery
Learnings
Guidance
Toolbox
Alignment
API Evangelist LLC

Hugging Face Audio API

Speech recognition, audio classification, and text-to-speech tasks

Hugging Face Audio API is one of 21 APIs that Hugging Face publishes on the APIs.io network, described by a machine-readable OpenAPI specification.

Tagged areas include Audio. The published artifact set on APIs.io includes an OpenAPI specification, API documentation, authentication docs, a getting-started guide, rate-limit docs, and pricing.

This API exposes 3 operations across 3 paths, and defines 1 schema. It is described by OpenAPI 3.1.0, at version 1.0.0.

Requests are made against a single base URL, https://datasets-server.huggingface.co.

3 operations 3 paths 1 schemas 3 POST

Metadata

The identity and technical contract details declared by the specification.

Specification
OpenAPI 3.1.0
API Version
1.0.0
Base URL
https://api-inference.huggingface.co
Authentication
HTTP Bearer
License
Terms of Service
Resource Areas
1

Authentication & Security 1

Hugging Face Audio API declares 1 security scheme for authenticating requests. It accepts HTTP bearer tokens (HF Token) (bearerAuth). By default, every request must be authenticated, though some operations may also be called without credentials.

  • bearerAuth — Optional Hugging Face API token. Required for private and gated datasets.

Paths & Operations 3

Across 3 paths, the API surfaces 3 operations — 3 POST. Each is listed below with its method, path, parameters, and response codes.

Audio 3

Speech recognition, audio classification, and text-to-speech tasks

POST
/models/{model_id}/automatic-speech-recognition
Automatic Speech Recognition Inference
automaticSpeechRecognition 1 param body → 200
POST
/v1/audio/transcriptions
Transcribe Audio
createTranscription body → 200
POST
/v1/audio/speech
Generate Speech
createSpeech body → 200

Schemas 1

The contract defines 1 schema that model the data the API accepts and returns. The most detailed is SpeechRecognitionResponse (1 property). Each schema is shown below with its type and property counts.

SpeechRecognitionResponse
object
1 property

Specification

The full machine-readable OpenAPI contract behind this narrative.

Source

hugging-face-audio-api-openapi.yml Raw ↑

Other APIs Hugging Face publishes across the network.

Hugging Face Chat API
Hugging Face Chat Completions API
Hugging Face Computer Vision API
Hugging Face Data Access API
Hugging Face Dataset Info API
Hugging Face Datasets API
Hugging Face Embeddings API
Hugging Face Endpoints API
Hugging Face Files & Metadata API
Hugging Face Image Generation API
Hugging Face Info API
Hugging Face Models API