How each transcript is written.
segments (default): transcript is plain text, and a separate segments field lists every caption line with its time, eg {"start": 18.64, "duration": 3.24, "text": "We're no strangers to love"}. In the results, open the Timed segments tab to see one row per line.
text: the whole transcript as one paragraph, no times. Good for AI, search and word counts.
timestamped: one line per caption with its time, eg [0:18] We're no strangers to love.
srt: a complete SRT subtitle file, ready to save as .srt.
vtt: a complete WebVTT subtitle file, ready to save as .vtt.
API example: "outputFormat": "timestamped"
Options:segmentstexttimestampedsrtvtt