API
POST /v1/audio/transcriptions
Upload audio as multipart form-data. Response text is billed in RODI.
POST
https://api.rodiumai.io/v1/audio/transcriptionsForm parameters
| Parameter | Type | Required | Description |
|---|---|---|---|
| file | file | Required | Audio file to transcribe (multipart form field). The whole request must stay under 10 MiB. |
| model | string | Required | Transcription model id (e.g. openai/gpt-4o-transcribe, openai/gpt-4o-mini-transcribe, google/gemini-2.5-flash). |
| language | string | Optional | Optional ISO-639-1 language hint (e.g. en, fr). |
| prompt | string | Optional | Optional text to guide style or spelling of the transcript. |
| response_format | string | Optional | Output format when supported (e.g. "json", "text", "verbose_json"). |
| chunking_strategy | string | Optional | OpenAI transcription models: chunking for long audio (e.g. "auto"). Forwarded as-is. |
| temperature | number | Optional | Sampling temperature between 0 and 1 when supported. |
Request examples
…SDK helpers
…Limits and behaviour
- Send
multipart/form-datawithfileandmodel. Missing fields return422with adetailarray instead of the usualerrorenvelope. - The whole request, file included, must stay under 10 MiB (
413 payload_too_largeabove). Split or compress longer recordings. - Both
Authorization: Bearerandx-api-keyare accepted. language,prompt,response_format,temperatureandchunking_strategyare forwarded to the model; the response is the provider's JSON (text, or richer fields withverbose_json).- Billed from the audio and text tokens the provider reports, at the model's price.