Duyu · STT
Transcription
Send an audio file, get the text. Same contract as OpenAI audio/transcriptions; your existing Whisper client works unchanged.
curl https://voice.patientdesk.ai/v1/audio/transcriptions \
-H "Authorization: Bearer pd_live_..." \
-F file=@kayit.m4a \
-F response_format=jsonForm fields
Send multipart/form-data. wav, mp3, m4a, webm, ogg and flac are accepted. Files longer than 30 s are split and re-joined automatically.
filefilerequired- The audio file. Up to 100 MB and 2 hours.
modelstringduyu-1.whisper-1is accepted for compatibility.- Default:
duyu-1 language"tr" | "auto"- The language spoken.
- Default:
tr promptstring- Spelling and vocabulary hint: drug names, person names, ID numbers. The model prefers these spellings.
timestamp_granularities[]string[]segmentand/orword; only withverbose_json.temperaturenumber- 0–1.
- Default:
0
Response
{ "text": "Merhaba, ben Ayşe Yılmaz. Sipariş numaram CNB123." }