multimodal

Speech to text

Transcribe audio to text through the Respan gateway with automatic logging.

post/api/audio/transcriptions

Headers

Authorizationstring required

Use your Respan API key for Respan API authentication. Enter only the Respan API key value; clients send Authorization: Bearer <RESPAN_API_KEY>. For /api/responses, OpenAI or Azure OpenAI provider credentials go in Settings -> Providers or the request body credential_override field, not in this auth field.

X-Data-Respan-Paramsstring

Base64-encoded JSON object of Respan parameters. Legacy X-Data-Keywordsai-Params is still accepted.

Response

Transcription result.

textstring

Transcribed text.

languagestring

Detected language.

durationnumber double

Audio duration in seconds.

segmentsApiAudioTranscriptionsPostResponsesContentApplicationJsonSchemaSegmentsItems[]

Segment-level timestamps (if requested).

Changes