curl --request POST \
--url https://api.vidgo.ai/api/generate/submit \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data @- <<EOF
{
"model": "elevenlabs/tts/turbo-v2.5",
"input": {
"similarity_boost": 0.75,
"style": 0,
"apply_text_normalization": "auto",
"text": "Almost there. Fold the top corner down until it meets the center crease. Now press along the edge with your fingertip. That's it. Turn the paper over, and we'll make the next fold together.",
"voice": "Rachel",
"language_code": "en",
"speed": 1.05,
"stability": 0.5,
"timestamps": false
}
}
EOF{
"code": 200,
"data": {
"task_id": "QKTPQMIP4AZY1LGE",
"status": "running",
"created_time": "2026-09-27T10:51:18"
}
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}"Request timeout"{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}ElevenLabs
ElevenLabs TTS Turbo v2.5
Create spoken narration with ElevenLabs Turbo v2.5, shaping pacing, vocal style, and voice similarity with optional text alignment.
POST
/
api
/
generate
/
submit
curl --request POST \
--url https://api.vidgo.ai/api/generate/submit \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data @- <<EOF
{
"model": "elevenlabs/tts/turbo-v2.5",
"input": {
"similarity_boost": 0.75,
"style": 0,
"apply_text_normalization": "auto",
"text": "Almost there. Fold the top corner down until it meets the center crease. Now press along the edge with your fingertip. That's it. Turn the paper over, and we'll make the next fold together.",
"voice": "Rachel",
"language_code": "en",
"speed": 1.05,
"stability": 0.5,
"timestamps": false
}
}
EOF{
"code": 200,
"data": {
"task_id": "QKTPQMIP4AZY1LGE",
"status": "running",
"created_time": "2026-09-27T10:51:18"
}
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}"Request timeout"{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}{
"detail": "<string>"
}- Submit the request with
VIDGO_API_KEY. Atask_idis returned immediately. Keep it until the task reachesfinishedorfailed. - Use Query Task Status to retrieve the result. If you provide a
callback_url, Vidgo sends the result to your webhook when the task finishes or fails.
ElevenLabs TTS Turbo v2.5
Create spoken narration with ElevenLabs Turbo v2.5, shaping pacing, vocal style, and voice similarity with optional text alignment.Available Models
- elevenlabs/tts/turbo-v2.5 - Convert text into speech.
Key Features
- Convert text into speech.
- Optionally request timestamps alongside the generated audio.
- Select a voice with
voiceand specify the spoken language withlanguage_code.
Advanced Parameters
Setmodel and optional callback_url at the request root. Place all workflow parameters, including output options, inside input.
The parameters below define the input object for this model.
Voice
voice: Optional. Select a voice name or compatible voice ID. String. Default:Rachel.
Stability
stability: Optional. Control voice stability. Number from0to1. Default:0.5.
Similarity Boost
similarity_boost: Optional. Control voice similarity. Number from0to1. Default:0.75.
Style
style: Optional. Set the voice style strength. Number from0to1. Default:0.
Speed
speed: Optional. Set the speech speed. Number from0.7to1.2. Default:1.
Timestamps
timestamps: Optional. Request timestamp information with the audio. Usetrueorfalse. Default:false.
Previous Text
previous_text: Optional. Supply preceding text as speech context. String. Also acceptsnull.
Next Text
next_text: Optional. Supply following text as speech context. String. Also acceptsnull.
Language Code
language_code: Optional. Set the spoken language with an ISO 639-1 code, such asen,fr,de,ja,vi,hu, orno. Also acceptsnull.
Apply Text Normalization
apply_text_normalization: Optional. Control how numbers, dates, and currencies are prepared for speech:autoselects normalization automatically,onenables it, andoffreads the supplied text directly. Default:auto.
Text
text: Required. Supply the text to speak. Length: 1-40,000 Unicode characters after trimming leading and trailing whitespace. Include at least one visible text character.
Output Files
- Read every item in the returned
filesarray. - Generated audio uses
file_type: "audio". - Timestamp and alignment files may use
file_type: "other". Preserve optional file metadata when present.
Examples
Each example pairs a request with its completed task response and generated audio. Submit the request toPOST https://api.vidgo.ai/api/generate/submit, then query GET https://api.vidgo.ai/api/generate/status/{task_id} with the task ID returned by your submission.
One fold at a time
One fold at a time
{
"model": "elevenlabs/tts/turbo-v2.5",
"input": {
"similarity_boost": 0.75,
"style": 0,
"apply_text_normalization": "auto",
"text": "Almost there. Fold the top corner down until it meets the center crease. Now press along the edge with your fingertip. That's it. Turn the paper over, and we'll make the next fold together.",
"voice": "Rachel",
"language_code": "en",
"speed": 1.05,
"stability": 0.5,
"timestamps": false
}
}
{
"code": 200,
"data": {
"task_id": "QKTPQMIP4AZY1LGE",
"status": "finished",
"files": [
{
"file_url": "https://cdn.vidgo.ai/apis/models/elevenlabs/tts/turbo-v2.5/v1/01/output.mp3",
"file_type": "audio"
}
],
"created_time": "2026-09-27T10:51:18",
"error_message": null,
"progress": 100
}
}
Finding north in the night sky
Finding north in the night sky
{
"model": "elevenlabs/tts/turbo-v2.5",
"input": {
"similarity_boost": 0.75,
"style": 0,
"apply_text_normalization": "auto",
"text": "今晚,我们先找到北斗七星,再从勺口两颗星的方向寻找北极星。它看起来并不耀眼,却能帮我们辨认北方。在黑暗中安静地等待片刻,你会发现,夜空远比第一眼看到的更丰富。",
"voice": "Aria",
"language_code": "zh",
"speed": 0.95,
"stability": 0.6,
"timestamps": true
}
}
{
"code": 200,
"data": {
"task_id": "2OAYYBTX2HQT9R8P",
"status": "finished",
"files": [
{
"file_url": "https://cdn.vidgo.ai/apis/models/elevenlabs/tts/turbo-v2.5/v1/02/output.mp3",
"file_type": "audio"
},
{
"file_url": "https://cdn.vidgo.ai/apis/models/elevenlabs/tts/turbo-v2.5/v1/02/attachment-1.json",
"file_type": "other"
}
],
"created_time": "2026-09-27T10:54:18",
"error_message": null,
"progress": 100
}
}
Two cups of tea
Two cups of tea
{
"model": "elevenlabs/tts/turbo-v2.5",
"input": {
"similarity_boost": 0.75,
"style": 0,
"apply_text_normalization": "auto",
"text": "Hoy volví a la casa de mi abuela. En la cocina seguía el reloj de siempre, cinco minutos adelantado. Preparé dos tazas de té por costumbre y, al darme cuenta, sonreí. Algunos gestos también guardan recuerdos.",
"voice": "Sarah",
"language_code": "es",
"speed": 0.98,
"stability": 0.45,
"timestamps": false,
"previous_text": "Hacía años que no cruzaba aquella puerta.",
"next_text": "Después abrí las ventanas para dejar entrar la luz."
}
}
{
"code": 200,
"data": {
"task_id": "KAQOQIL83QO5TWA6",
"status": "finished",
"files": [
{
"file_url": "https://cdn.vidgo.ai/apis/models/elevenlabs/tts/turbo-v2.5/v1/03/output.mp3",
"file_type": "audio"
}
],
"created_time": "2026-09-27T10:52:07",
"error_message": null,
"progress": 100
}
}
Authorizations
Use VIDGO_API_KEY.
Body
application/json
Place model and optional callback_url at the request root, and the speech parameters inside input.