Fields
| Field | Type | Required | Description |
|---|---|---|---|
File | operations.CreateAudioTranscriptionsMultipartFile | :heavy_check_mark: | The audio file to transcribe. The format is derived from the filename extension or the file part content type. Max 25 MB; send larger files as base64 JSON via input_audio. |
Language | *string | :heavy_minus_sign: | The language of the input audio (ISO-639-1). |
Model | string | :heavy_check_mark: | The model to use for transcription. |
ResponseFormat | *operations.ResponseFormat | :heavy_minus_sign: | The response format. βjsonβ (default) returns { text, usage }; βverbose_jsonβ additionally returns task, language, duration, and segment-level timestamps (OpenAI-compatible providers only). |
Temperature | *float64 | :heavy_minus_sign: | The sampling temperature. |
TimestampGranularities | []operations.TimestampGranularities | :heavy_minus_sign: | Timestamp detail levels to include when response_format is βverbose_jsonβ. βwordβ additionally returns word-level timestamps in the words array. |