Speech to text
Authentication
Use your Respan API key for Respan API authentication. Enter only the Respan API key value; clients send Authorization: Bearer <RESPAN_API_KEY>. For /api/responses, provider credentials such as Perplexity, OpenAI, or Azure OpenAI go in Settings -> Providers or respan_params.credential_override in the request body, not in this authentication field.
Headers
Base64-encoded JSON object of Respan parameters. Legacy X-Data-Keywordsai-Params is still accepted.
Request
Audio file. Supported: flac, mp3, mp4, mpeg, mpga, m4a, ogg, wav, webm.
Input audio language (ISO-639-1).
Sampling temperature (0-1).
Timestamp granularities. Requires verbose_json response format.
Per-customer LLM provider credentials.
When true, omits input/output from the log. Metrics still recorded.
Custom key-value metadata.
Response
Word-level timestamps (if requested).
Segment-level timestamps (if requested).