Eleven v4
NewPopularElevenLabs · Audio · Text → Speech
Most expressive voice engine — audio tags direct the delivery, 90+ languages.
eleven-v4Workflow: elevenlabs/v1/text-to-speechOverview
Alongside the unified model APIs, we expose compatibility APIs that take each vendor's original parameters exactly as the vendor defines them — nothing renamed, nothing reshaped on the way through.
That makes them the shortest path onto Picsart if you already work with the vendor directly: the request bodies you've already written keep working as they are. You keep their parameter names and defaults, and you get the vendor's full parameter set rather than the subset that is shared across every model.
The cost is that a request is written for one vendor — switching models later means rewriting it, and results come back in the vendor's own shape. When you'd rather write once and change models freely, use the unified API.
Make a request
Call the workflow with ai.apis.run() in TypeScript, or hit /workflows/{workflow}/execute directly — params is passed through untouched either way.
These endpoints are addressed by workflow name rather than model id: the one in this model's header, since one model runs as one workflow. Authentication is unchanged — your Picsart API key as a bearer token (see Authentication).
import { createClient, ApiRunMode } from '@picsart/ai-sdk';
const ai = createClient({
apiKey: process.env.PICSART_API_KEY,
apiUrl: 'https://api.picsart.com',
});
// Calls the 'elevenlabs/v1/text-to-speech' workflow directly — params are sent as-is.
const { result, usage } = await ai.apis.run('elevenlabs/v1/text-to-speech', {
voice_id: "21m00Tcm4TlvDq8ikWAM",
text: "Hello, welcome to our product demo."
}, {
mode: ApiRunMode.SYNC,
});
console.log(result); // workflow-specific output
console.log(usage?.credits); // credits chargedAsync (submit & poll)
This model can run longer than the sync limit (~20s). In TypeScript, ai.apis.run(…, { mode: ApiRunMode.ASYNC }) polls for you; over HTTP, submit and poll yourself.
import { createClient, ApiRunMode } from '@picsart/ai-sdk';
const ai = createClient({
apiKey: process.env.PICSART_API_KEY,
apiUrl: 'https://api.picsart.com',
});
// mode: ASYNC submits the job and polls under the hood — you just await.
const { result } = await ai.apis.run('elevenlabs/v1/text-to-speech', {
voice_id: "21m00Tcm4TlvDq8ikWAM",
text: "Hello, welcome to our product demo."
}, {
mode: ApiRunMode.ASYNC,
});
console.log(result);Parameters
9 parameters, sent inside params. These are the vendor's own names, so they line up one-for-one with the vendor's documentation. Required ones must be supplied; the rest fall back to their defaults.
| Parameter | Type | Required | Default | Details |
|---|---|---|---|---|
voice_idID of the ElevenLabs voice to use | string | yes | — | e.g. 21m00Tcm4TlvDq8ikWAM |
textThe text to convert to speech. On eleven_v3, eleven_v4 and eleven_v4_turbo, bracketed audio tags such as [whispering] or [laughing] direct the delivery. Per-request cap: 10,000 characters (5,000 on eleven_v3). | string | yes | — | e.g. Hello, welcome to our product demo. |
model_idThe model to use for text-to-speech. eleven_v4 is the most expressive engine (90+ languages, 10,000 characters); eleven_v4_turbo trades some quality for latency. With a language_code that differs from the voice, v4 speaks the target language natively instead of carrying the voice's accent. | string | no | eleven_multilingual_v2 | eleven_multilingual_v2eleven_v3eleven_v4eleven_v4_turbo |
output_formatOutput audio format | string | no | mp3_44100_128 | mp3_22050_32mp3_44100_32mp3_44100_64mp3_44100_96mp3_44100_128mp3_44100_192mp3_48000_192pcm_16000pcm_22050pcm_24000pcm_44100pcm_48000 |
voice_settingsVoice settings to override stored defaults | object | no | — | — |
└ stabilityVoice consistency (0-1). Lower = more variation. | number | no | — | — |
└ similarity_boostAdherence to original voice (0-1) | number | no | — | — |
└ styleStyle exaggeration intensity (0-1). Ignored by eleven_v4 and eleven_v4_turbo, which take delivery direction from audio tags in the text instead. | number | no | — | — |
└ speedSpeech rate multiplier. Ignored by eleven_v4 and eleven_v4_turbo. | number | no | 1 | — |
└ use_speaker_boostBoost similarity to the original speaker. Ignored by eleven_v4 and eleven_v4_turbo. | boolean | no | true | — |
language_codeISO 639-1 language code for text normalization | string | no | — | e.g. en |
with_timestampsReturn per-character timings alongside the audio, for captions and lip sync | boolean | no | false | — |
seedSeed for deterministic generation (0-4294967295) | number | no | — | — |
optionsOptions controlling safety checks and drive integration | object | no | — | — |
└ safety_checksSafety check settings | object | no | — | — |
└ enabledWhether to run content moderation. Defaults to true. | boolean | no | — | — |
└ driveSave result to Picsart Drive | object | no | — | — |
└ nameFile name in Picsart Drive | string | yes | — | — |
└ attributesCustom attributes to attach to the file | object | no | — | — |
└ folderTarget folder in Picsart Drive | object | no | — | — |
└ inputs_transformationInput transformation settings | object | no | — | — |
└ downscale_oversized_imagesWhether to downscale oversized input images. Defaults to false. | boolean | no | — | — |
Response
Over HTTP the output arrives inside a status envelope, at response.result. ai.apis.run() unwraps that envelope for you and resolves to { result, usage } instead. Either way the resultitself is the vendor's own shape.
{
"result": {
"url": "https://cdn.picsart.com/…/result.mp3",
"mimeType": "audio/mpeg",
"driveFile": {},
"alignment": {
"characters": [
"…"
],
"character_start_times_seconds": [
0
],
"character_end_times_seconds": [
0
]
},
"normalizedAlignment": {
"characters": [
"…"
],
"character_start_times_seconds": [
0
],
"character_end_times_seconds": [
0
]
}
},
"usage": {
"credits": 1
}
}