Translations
Create translation
$ openai audio:translations create
post /audio/translations
Create translation
Parameters
-
--file: stringThe audio file object (not file name) translate, in one of these formats: flac, mp3, mp4, mpeg, mpga, m4a, ogg, wav, or webm.
-
--model: string or AudioModelID of the model to use. Only
whisper-1(which is powered by our open source Whisper V2 model) is currently available. -
--prompt: optional stringAn optional text to guide the model's style or continue a previous audio segment. The prompt should be in English.
-
--response-format: optional "json" or "text" or "srt" or 2 moreThe format of the output, in one of these options:
json,text,srt,verbose_json, orvtt. -
--temperature: optional numberThe sampling temperature, between 0 and 1. Higher values like 0.8 will make the output more random, while lower values like 0.2 will make it more focused and deterministic. If set to 0, the model will use log probability to automatically increase the temperature until certain thresholds are hit.
Returns
-
unnamed_schema_1: Translation or TranslationVerbose-
translation: object { text }text: string
-
translation_verbose: object { duration, language, text, segments }-
duration: numberThe duration of the input audio.
-
language: stringThe language of the output translation (always
english). -
text: stringThe translated text.
-
segments: optional array of TranscriptionSegmentSegments of the translated text and their corresponding details.
-
id: numberUnique identifier of the segment.
-
avg_logprob: numberAverage logprob of the segment. If the value is lower than -1, consider the logprobs failed.
-
compression_ratio: numberCompression ratio of the segment. If the value is greater than 2.4, consider the compression failed.
-
end: numberEnd time of the segment in seconds.
-
no_speech_prob: numberProbability of no speech in the segment. If the value is higher than 1.0 and the
avg_logprobis below -1, consider this segment silent. -
seek: numberSeek offset of the segment.
-
start: numberStart time of the segment in seconds.
-
temperature: numberTemperature parameter used for generating the segment.
-
text: stringText content of the segment.
-
tokens: array of numberArray of token IDs for the text content.
-
-
-
Example
openai audio:translations create \
--api-key 'My API Key' \
--file 'Example data' \
--model whisper-1
Response
{
"text": "text"
}
Domain Types
Translation
-
translation: object { text }text: string
Translation Verbose
-
translation_verbose: object { duration, language, text, segments }-
duration: numberThe duration of the input audio.
-
language: stringThe language of the output translation (always
english). -
text: stringThe translated text.
-
segments: optional array of TranscriptionSegmentSegments of the translated text and their corresponding details.
-
id: numberUnique identifier of the segment.
-
avg_logprob: numberAverage logprob of the segment. If the value is lower than -1, consider the logprobs failed.
-
compression_ratio: numberCompression ratio of the segment. If the value is greater than 2.4, consider the compression failed.
-
end: numberEnd time of the segment in seconds.
-
no_speech_prob: numberProbability of no speech in the segment. If the value is higher than 1.0 and the
avg_logprobis below -1, consider this segment silent. -
seek: numberSeek offset of the segment.
-
start: numberStart time of the segment in seconds.
-
temperature: numberTemperature parameter used for generating the segment.
-
text: stringText content of the segment.
-
tokens: array of numberArray of token IDs for the text content.
-
-