Skip to main content
POST
  1. Submit the request with VIDGO_API_KEY. A task_id is returned immediately. Keep it until the task reaches finished or failed.
  2. Use Query Task Status to retrieve the result. If you provide a callback_url, Vidgo sends the result to your webhook when the task finishes or fails.

ElevenLabs TTS Turbo v2.5

Create spoken narration with ElevenLabs Turbo v2.5, shaping pacing, vocal style, and voice similarity with optional text alignment.

Available Models

  • elevenlabs/tts/turbo-v2.5 - Convert text into speech.

Key Features

  • Convert text into speech.
  • Optionally request timestamps alongside the generated audio.
  • Select a voice with voice and specify the spoken language with language_code.

Advanced Parameters

Set model and optional callback_url at the request root. Place all workflow parameters, including output options, inside input. The parameters below define the input object for this model.

Voice

  • voice: Optional. Select a voice name or compatible voice ID. String. Default: Rachel.

Stability

  • stability: Optional. Control voice stability. Number from 0 to 1. Default: 0.5.

Similarity Boost

  • similarity_boost: Optional. Control voice similarity. Number from 0 to 1. Default: 0.75.

Style

  • style: Optional. Set the voice style strength. Number from 0 to 1. Default: 0.

Speed

  • speed: Optional. Set the speech speed. Number from 0.7 to 1.2. Default: 1.

Timestamps

  • timestamps: Optional. Request timestamp information with the audio. Use true or false. Default: false.

Previous Text

  • previous_text: Optional. Supply preceding text as speech context. String. Also accepts null.

Next Text

  • next_text: Optional. Supply following text as speech context. String. Also accepts null.

Language Code

  • language_code: Optional. Set the spoken language with an ISO 639-1 code, such as en, fr, de, ja, vi, hu, or no. Also accepts null.

Apply Text Normalization

  • apply_text_normalization: Optional. Control how numbers, dates, and currencies are prepared for speech: auto selects normalization automatically, on enables it, and off reads the supplied text directly. Default: auto.

Text

  • text: Required. Supply the text to speak. Length: 1-40,000 Unicode characters after trimming leading and trailing whitespace. Include at least one visible text character.

Output Files

  • Read every item in the returned files array.
  • Generated audio uses file_type: "audio".
  • Timestamp and alignment files may use file_type: "other". Preserve optional file metadata when present.

Examples

Each example pairs a request with its completed task response and generated audio. Submit the request to POST https://api.vidgo.ai/api/generate/submit, then query GET https://api.vidgo.ai/api/generate/status/{task_id} with the task ID returned by your submission.
Download audio
Download audioDownload attachment-1.json
Download audio

Authorizations

Authorization
string
header
required

Use VIDGO_API_KEY.

Body

application/json

Place model and optional callback_url at the request root, and the speech parameters inside input.

model
string
required
Example:

"elevenlabs/tts/turbo-v2.5"

input
object
required
callback_url
string | null

Optional public HTTP(S) endpoint for task completion notifications. HTTPS is recommended. Omit, use null, or use an empty string to disable callbacks.

Response

Task submitted

code
enum<integer>
required
Available options:
200
data
object
required