{"id":260798,"date":"2026-07-30T16:45:00","date_gmt":"2026-07-30T23:45:00","guid":{"rendered":"https:\/\/picsart.com\/blog\/?p=260798"},"modified":"2026-07-30T16:51:20","modified_gmt":"2026-07-30T23:51:20","slug":"how-to-use-elevenlabs-text-to-speech","status":"publish","type":"post","link":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/","title":{"rendered":"How to use ElevenLabs text to speech in Picsart"},"content":{"rendered":"<p>ElevenLabs text to speech turns a written script into studio-quality spoken audio, and the model you pick decides what kind of performance you get. Eleven v3 acts a script with emotion and direction. Eleven Multilingual v2 narrates it cleanly and holds steady across long recordings. Both cost the same, so the choice is about the read rather than the budget.<\/p>\n<p>This guide covers what each model is built for, how to direct the delivery with audio tags, how to generate a two-person conversation, and how to run the whole thing in a browser.<\/p>\n<h2><span id=\"The_two_models_and_what_each_one_is_built_for\">The two models and what each one is built for<\/span><\/h2>\n<p><strong>Eleven v3<\/strong> is the expressive one. It covers more than 70 languages, carries audio tags and multi-speaker dialogue, takes 5,000 characters per generation for about five minutes of speech, and is still marked experimental.<\/p>\n<p><strong>Eleven Multilingual v2<\/strong> is the stable, professional option. It covers 29 languages, takes 10,000 characters for about ten minutes in one pass, and is what ElevenLabs themselves recommend for content creation, audiobooks and video narration.<\/p>\n<p>Both run at 3 credits per 1,000 characters, so the choice is about the read rather than the cost.<\/p>\n<p>That surprises people. Eleven v3 is newer and more capable, but for a straight voiceover on a finished video, consistency across a long read matters more than dramatic range, and Multilingual v2 also gives you double the script length in one pass.<\/p>\n<h2><span id=\"The_model_reads_the_room_not_just_the_words\">The model reads the room, not just the words<\/span><\/h2>\n<p>Before you write a single instruction, the voice is already responding to what your script says. ElevenLabs voices pick up emotional cues in the text and adapt delivery to both the sentence at hand and the wider context around it, which is what stops a line landing in the wrong tone when the subject changes.<\/p>\n<p>That matters in practice, because a well-written script often needs no direction at all. Tags are for the moments where you want something the words alone would not produce.<\/p>\n<h2><span id=\"Directing_delivery_with_audio_tags\">Directing delivery with audio tags<\/span><\/h2>\n<p>Audio tags are what set Eleven v3 apart. Write a cue in square brackets, drop it inline in the script, and the delivery changes at exactly that point.<\/p>\n<pre><code>[slowly] Back then... [chuckles] we had no phones.\r\n[whispers] Just dirt roads and [coughs] big dreams. [sad] Then it happened.<\/code><\/pre>\n<p>The tags group into three kinds:<\/p>\n<ul>\n<li><strong>Emotion.<\/strong> [sad], [angry], [curious], [sarcastically], [mischievously]<\/li>\n<li><strong>Delivery and pace.<\/strong> [whispers], [shouts], [slowly]<\/li>\n<li><strong>Human reactions.<\/strong> [laughs], [chuckles], [giggles], [sighs], [clears throat], [coughs]<\/li>\n<\/ul>\n<p>Capital letters carry weight too, so a shouted line reads as <strong>And GOOOOAL!<\/strong> rather than leaning on a tag alone. Character colour comes from stacked cues like [laughs wickedly] and [evil laugh].<\/p>\n<p>Longer tag lists circulate online, but they are community-collected and not all of them fire. Sticking to the documented set is the difference between a script that performs and one that reads your notes aloud.<\/p>\n<h2><span id=\"Writing_a_script_the_model_reads_well\">Writing a script the model reads well<\/span><\/h2>\n<p>Punctuation is performance direction. Commas and full stops create the pauses, so a run-on sentence gets read as a run-on sentence. Spell out anything unusual the way you want it said, including acronyms, product names and numbers, and respell a stubborn word phonetically if it keeps coming out wrong.<\/p>\n<p>Give the model enough runway as well. A clipped fragment offers no context for setting tone, so a few full sentences almost always read better than one short line.<\/p>\n<h2><span id=\"Generating_a_two-person_conversation\">Generating a two-person conversation<\/span><\/h2>\n<p>Eleven v3 supports natural multi-speaker dialogue. Label each speaker on its own line and let a single generation carry the whole exchange:<\/p>\n<pre><code>Mark\r\nHey Chris... Knock knock.\r\n\r\nChris\r\n[chuckles] I'm not doing this AGAIN!\r\n\r\nMark\r\n[laughing] Come on, PLEASE! I promise you'll love this one.<\/code><\/pre>\n<p>Speakers in one generation share context, so timing, interruptions and reactions all land naturally, and even overlapping speech survives. Generating each part separately and stitching the takes together loses precisely that, and the joins are audible.<\/p>\n<h2><span id=\"Choosing_a_voice\">Choosing a voice<\/span><\/h2>\n<p>Voices are grouped by the job they are built for, which is a faster way in than scrolling a list. Narration voices carry audiobooks and podcasts. Conversational voices suit informal, spoken-word scenarios. Character voices are built for cartoons and games. Social media voices are made to hold attention in short-form video, and advertisement voices aim at recall and action.<\/p>\n<h2><span id=\"What_the_finished_audio_sounds_like\">What the finished audio sounds like<\/span><\/h2>\n<p>Speech exports as MP3 at 44.1 kHz and 128 kbps, which is the same quality you would expect from a music file and more than enough for social video, podcasts and voiceover work. Once it is generated, drop it into the <a href=\"https:\/\/picsart.com\/video-editor\/\">video editor<\/a> to lay it over your footage.<\/p>\n<h2><span id=\"When_output_goes_wrong\">When output goes wrong<\/span><\/h2>\n<figure class=\"wp-block-table\">\n<table style=\"border-collapse: collapse; width: 100%; table-layout: auto;\">\n<thead>\n<tr>\n<th style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #000000; font-weight: bold; white-space: nowrap;\">Problem<\/th>\n<th style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #000000; font-weight: bold; white-space: nowrap;\">Fix<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Audio tags ignored<\/td>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Switch to Eleven v3<\/td>\n<\/tr>\n<tr>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Word mispronounced<\/td>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Respell it phonetically<\/td>\n<\/tr>\n<tr>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Read is too flat<\/td>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Add a delivery tag, give the line more context<\/td>\n<\/tr>\n<tr>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Voice shifts between parts<\/td>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Same model and voice throughout, split at paragraph ends<\/td>\n<\/tr>\n<tr>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Quality drifts on long audio<\/td>\n<td style=\"border: 1px solid #333333; padding: 10px 14px; text-align: left; vertical-align: top; color: #ffffff; background: #141414;\">Move to Multilingual v2<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/figure>\n<h2><span id=\"The_rest_of_the_ElevenLabs_audio_models\">The rest of the ElevenLabs audio models<\/span><\/h2>\n<p>Text to speech is one part of a wider set. <strong>Eleven Dubbing<\/strong> is the standout, because it is free and it takes a finished video into another language with the voices matched. <strong>Eleven STS v2<\/strong> swaps one voice for another while keeping the original timing, and <strong>Eleven Audio Isolation<\/strong> strips background noise from a messy recording. For audio that does not exist yet, <strong>ElevenLabs SFX v2<\/strong> makes sound effects up to 15 seconds long, the AI music generator runs <strong>Music v2<\/strong> for full tracks, and <strong>Eleven Voice Design v3<\/strong> builds a brand new voice from a description.<\/p>\n<p>Generating a voiceover takes a script, a model and a voice.<\/p>\n<section class=\"section_how_to\">\n            <h2 class=\"how_to_title\" id=\"How_to_use_ElevenLabs_text_to_speech_with_Picsart\">How to use ElevenLabs text to speech with Picsart<\/h2>\n    \n        <div class=\"how_to_steps\">\n                                        <div class=\"how_to_step how_to_step--highlighted\">\n                    <div class=\"how_to_step_content\">\n                        <p class=\"how_to_step_title\">\n                            <span class=\"how_to_step_number\">1.<\/span>\n                            Open the AI Playground                        <\/p>\n                                                    <p class=\"how_to_step_description\">Go to the Picsart AI Playground and switch to Audio mode. Everything runs in the browser, so there is nothing to install and nothing to set up.<\/p>\n                                                                            <div class=\"how_to_cta_wrapper\">\n                                                                <input\n                                    type=\"file\"\n                                    id=\"how_to_upload_how-to-6a6c146c78208_0\"\n                                    class=\"how_to_upload_input\"\n                                    accept=\"image\/*\"\n                                    data-deeplink=\"https:\/\/picsart.com\/ai-playground\/\"\n                                \/>\n                                <button\n                                    type=\"button\"\n                                    class=\"how_to_cta_button\"\n                                    data-upload-id=\"how_to_upload_how-to-6a6c146c78208_0\"\n                                >\n                                    <img loading=\"lazy\" decoding=\"async\" src=\"https:\/\/cdn-cms-uploads.picsart.com\/cms-uploads\/9b784b6b-6f78-4ee4-a748-f4ad781bfd34.svg\" alt=\"\" width=\"20\" height=\"20\" class=\"how_to_cta_icon\" \/>\n                                    <span>Try it in the AI Playground<\/span>\n                                <\/button>\n                                                            <\/div>\n                                            <\/div>\n                <\/div>\n                                                <div class=\"how_to_step how_to_step--highlighted\">\n                    <div class=\"how_to_step_content\">\n                        <p class=\"how_to_step_title\">\n                            <span class=\"how_to_step_number\">2.<\/span>\n                            Choose your model                        <\/p>\n                                                    <p class=\"how_to_step_description\">Pick Eleven v3 for an expressive, performed read, or Eleven Multilingual v2 for steady narration and a longer script in one pass. Auto Mode will otherwise choose a model for you.<\/p>\n                                                                    <\/div>\n                <\/div>\n                                                <div class=\"how_to_step how_to_step--highlighted\">\n                    <div class=\"how_to_step_content\">\n                        <p class=\"how_to_step_title\">\n                            <span class=\"how_to_step_number\">3.<\/span>\n                            Paste your script                        <\/p>\n                                                    <p class=\"how_to_step_description\">Add the words you want spoken. Punctuate deliberately, because commas and full stops create the pauses, and spell out acronyms and numbers the way you want them said. On Eleven v3 you can write audio tags inline in square brackets.<\/p>\n                                                                    <\/div>\n                <\/div>\n                                                <div class=\"how_to_step how_to_step--highlighted\">\n                    <div class=\"how_to_step_content\">\n                        <p class=\"how_to_step_title\">\n                            <span class=\"how_to_step_number\">4.<\/span>\n                            Set the language and accent                        <\/p>\n                                                    <p class=\"how_to_step_description\">Both are plain text fields rather than menus, so describe what you need. Naming a regional accent here is what produces one.<\/p>\n                                                                    <\/div>\n                <\/div>\n                                                <div class=\"how_to_step how_to_step--highlighted\">\n                    <div class=\"how_to_step_content\">\n                        <p class=\"how_to_step_title\">\n                            <span class=\"how_to_step_number\">5.<\/span>\n                            Pick a voice                        <\/p>\n                                                    <p class=\"how_to_step_description\">Audition two or three voices on a couple of sentences of your real script rather than on a sample line. A voice that fumbles a product name will keep fumbling it, and that only shows up on your own copy.<\/p>\n                                                                    <\/div>\n                <\/div>\n                                                <div class=\"how_to_step how_to_step--highlighted\">\n                    <div class=\"how_to_step_content\">\n                        <p class=\"how_to_step_title\">\n                            <span class=\"how_to_step_number\">6.<\/span>\n                            Generate and save                        <\/p>\n                                                    <p class=\"how_to_step_description\">Select Generate to see the credit cost, then save the audio to a project board so you can reuse it across designs and videos.<\/p>\n                                                                    <\/div>\n                <\/div>\n                        <\/div>\n    <\/section>\n\n<script type=\"application\/ld+json\">\n{\n    \"@context\": \"https:\/\/schema.org\",\n    \"@type\": \"HowTo\",\n    \"name\": \"How to use ElevenLabs text to speech with Picsart\",\n    \"step\": [\n        {\n            \"@type\": \"HowToStep\",\n            \"position\": 1,\n            \"name\": \"Open the AI Playground\",\n            \"text\": \"Go to the Picsart AI Playground and switch to Audio mode. Everything runs in the browser, so there is nothing to install and nothing to set up.\"\n        },\n        {\n            \"@type\": \"HowToStep\",\n            \"position\": 2,\n            \"name\": \"Choose your model\",\n            \"text\": \"Pick Eleven v3 for an expressive, performed read, or Eleven Multilingual v2 for steady narration and a longer script in one pass. Auto Mode will otherwise choose a model for you.\"\n        },\n        {\n            \"@type\": \"HowToStep\",\n            \"position\": 3,\n            \"name\": \"Paste your script\",\n            \"text\": \"Add the words you want spoken. Punctuate deliberately, because commas and full stops create the pauses, and spell out acronyms and numbers the way you want them said. On Eleven v3 you can write audio tags inline in square brackets.\"\n        },\n        {\n            \"@type\": \"HowToStep\",\n            \"position\": 4,\n            \"name\": \"Set the language and accent\",\n            \"text\": \"Both are plain text fields rather than menus, so describe what you need. Naming a regional accent here is what produces one.\"\n        },\n        {\n            \"@type\": \"HowToStep\",\n            \"position\": 5,\n            \"name\": \"Pick a voice\",\n            \"text\": \"Audition two or three voices on a couple of sentences of your real script rather than on a sample line. A voice that fumbles a product name will keep fumbling it, and that only shows up on your own copy.\"\n        },\n        {\n            \"@type\": \"HowToStep\",\n            \"position\": 6,\n            \"name\": \"Generate and save\",\n            \"text\": \"Select Generate to see the credit cost, then save the audio to a project board so you can reuse it across designs and videos.\"\n        }\n    ]\n}<\/script>\n\n<script>\n(function() {\n    function uploadFallback(file) {\n        var UPLOAD_URL = 'https:\/\/upload.picsart.com\/files';\n        var UPLOAD_URL_STAGE2 = 'https:\/\/upload-stage.picsartstage2.com\/files';\n        var hostname = window.location.hostname;\n        var isStage2 = hostname.indexOf('picsartstage2.com') !== -1 || hostname.indexOf('stage2') !== -1;\n        var url = isStage2 ? UPLOAD_URL_STAGE2 : UPLOAD_URL;\n        var isSafari = \/Safari\/i.test(navigator.userAgent) && !\/Chrome|Chromium|FxiOS|Edg|OPR\/i.test(navigator.userAgent);\n        if (isSafari) {\n            return new Promise(function(resolve, reject) {\n                var formData = new FormData();\n                formData.append('type', 'editing-temp-landings');\n                formData.append('file', file);\n                formData.append('url', '');\n                formData.append('metainfo', '');\n                var xhr = new XMLHttpRequest();\n                xhr.open('POST', url);\n                xhr.onload = function() {\n                    try {\n                        var data = xhr.responseText ? JSON.parse(xhr.responseText) : null;\n                        if (xhr.status >= 200 && xhr.status < 300 && data && data.result && data.result.url) {\n                            resolve(data.result.url);\n                        } else {\n                            reject(new Error('Upload failed'));\n                        }\n                    } catch (e) { reject(new Error('Upload failed')); }\n                };\n                xhr.onerror = xhr.ontimeout = function() { reject(new Error('Upload failed')); };\n                xhr.timeout = 60000;\n                xhr.send(formData);\n            });\n        }\n        var formData = new FormData();\n        formData.append('type', 'editing-temp-landings');\n        formData.append('file', file);\n        formData.append('url', '');\n        formData.append('metainfo', '');\n        return fetch(url, { method: 'POST', body: formData, mode: 'cors', cache: 'no-store' })\n            .then(function(res) { return res.text(); })\n            .then(function(text) {\n                try {\n                    var data = text ? JSON.parse(text) : null;\n                    if (data && data.result && data.result.url) return data.result.url;\n                } catch (e) { }\n                throw new Error('Upload failed');\n            });\n    }\n    var uploadFileToCDN = window.HowToUpload && window.HowToUpload.uploadFileToCDN\n        ? window.HowToUpload.uploadFileToCDN\n        : uploadFallback;\n\n    if (window._howToUploadBound) return;\n    window._howToUploadBound = true;\n\n    document.addEventListener('click', function(e) {\n        var btn = e.target && e.target.closest && e.target.closest('.how_to_cta_button');\n        if (!btn) return;\n        var uploadId = btn.getAttribute('data-upload-id');\n        var input = uploadId ? document.getElementById(uploadId) : null;\n        if (input) {\n            e.preventDefault();\n            input.click();\n        }\n    }, true);\n\n    document.addEventListener('change', function(e) {\n        if (!e.target || !e.target.classList || !e.target.classList.contains('how_to_upload_input')) return;\n        var file = e.target.files && e.target.files[0];\n        if (!file) return;\n        var inputEl = e.target;\n        var deeplink = inputEl.getAttribute('data-deeplink');\n        var button = document.querySelector('.how_to_cta_button[data-upload-id=\"' + inputEl.id + '\"]');\n        var labelSpan = button ? button.querySelector('span') : null;\n        var originalLabelText = labelSpan ? labelSpan.textContent : '';\n        if (button) {\n            button.disabled = true;\n            if (labelSpan) labelSpan.textContent = 'Uploading\u2026';\n        }\n        uploadFileToCDN(file)\n            .then(function(cdnUrl) {\n                var separator = deeplink.indexOf('?') !== -1 ? '&' : '?';\n                var params = 'ref=blog&image=' + encodeURIComponent(cdnUrl);\n                var redirectUrl = deeplink + separator + params;\n                inputEl.value = '';\n                setTimeout(function() {\n                    window.location.assign(redirectUrl);\n                }, 0);\n            })\n            .catch(function() {\n                if (button) button.disabled = false;\n                if (labelSpan) labelSpan.textContent = originalLabelText;\n                inputEl.value = '';\n                alert('Upload failed. Please try again.');\n            });\n    }, true);\n})();\n<\/script>\n\n<section class=\"section_faq\" id=\"faq-faq-6a6c146c784e6\">\n            <h2 class=\"faq_title\" id=\"Get_answers_to_common_questions\">Get answers to common questions<\/h2>\n    \n    <div class=\"faq_items\">\n                    <div class=\"faq_item faq_item--active\">\n                <button type=\"button\" class=\"faq_question\" aria-expanded=\"true\">\n                    <span class=\"faq_question_text\">Which ElevenLabs model should I use for text to speech?<\/span>\n                    <svg class=\"faq_chevron\" width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                        <path d=\"M6 9L12 15L18 9\" stroke=\"currentColor\" stroke-width=\"1.5\" stroke-linecap=\"round\" stroke-linejoin=\"round\"\/>\n                    <\/svg>\n                <\/button>\n                <div class=\"faq_answer\" aria-hidden=\"false\">\n                    <div class=\"faq_answer_content\"><p>Eleven v3 for storytelling, character work and anything that needs a performed read, since it is the model that responds to audio tags. Eleven Multilingual v2 for voiceovers, audiobooks and long recordings that need a consistent voice. Both cost the same in Picsart, so the choice is about the read rather than the budget.<\/p>\n<\/div>\n                <\/div>\n                <div class=\"faq_divider\"><\/div>\n            <\/div>\n                    <div class=\"faq_item \">\n                <button type=\"button\" class=\"faq_question\" aria-expanded=\"false\">\n                    <span class=\"faq_question_text\">How do audio tags work in ElevenLabs?<\/span>\n                    <svg class=\"faq_chevron\" width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                        <path d=\"M6 9L12 15L18 9\" stroke=\"currentColor\" stroke-width=\"1.5\" stroke-linecap=\"round\" stroke-linejoin=\"round\"\/>\n                    <\/svg>\n                <\/button>\n                <div class=\"faq_answer\" aria-hidden=\"true\" data-collapsed>\n                    <div class=\"faq_answer_content\"><p>Write them inline in your script inside square brackets, such as [whispers], [laughs] or [slowly], placed exactly where the delivery should change. Eleven v3 handles the full range of emotion, direction and audio effects. Multilingual v2 handles basic pauses and breaks only.<\/p>\n<\/div>\n                <\/div>\n                <div class=\"faq_divider\"><\/div>\n            <\/div>\n                    <div class=\"faq_item \">\n                <button type=\"button\" class=\"faq_question\" aria-expanded=\"false\">\n                    <span class=\"faq_question_text\">What is the ElevenLabs character limit?<\/span>\n                    <svg class=\"faq_chevron\" width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                        <path d=\"M6 9L12 15L18 9\" stroke=\"currentColor\" stroke-width=\"1.5\" stroke-linecap=\"round\" stroke-linejoin=\"round\"\/>\n                    <\/svg>\n                <\/button>\n                <div class=\"faq_answer\" aria-hidden=\"true\" data-collapsed>\n                    <div class=\"faq_answer_content\"><p>Eleven v3 takes 5,000 characters per generation, roughly five minutes of speech. Eleven Multilingual v2 takes 10,000, roughly ten minutes. For longer content, split the script across several generations.<\/p>\n<\/div>\n                <\/div>\n                <div class=\"faq_divider\"><\/div>\n            <\/div>\n                    <div class=\"faq_item \">\n                <button type=\"button\" class=\"faq_question\" aria-expanded=\"false\">\n                    <span class=\"faq_question_text\">Can ElevenLabs generate a conversation between two speakers?<\/span>\n                    <svg class=\"faq_chevron\" width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                        <path d=\"M6 9L12 15L18 9\" stroke=\"currentColor\" stroke-width=\"1.5\" stroke-linecap=\"round\" stroke-linejoin=\"round\"\/>\n                    <\/svg>\n                <\/button>\n                <div class=\"faq_answer\" aria-hidden=\"true\" data-collapsed>\n                    <div class=\"faq_answer_content\"><p>Yes. Eleven v3 supports natural multi-speaker dialogue, where the speakers share context inside a single generation so they react to one another and the timing sounds natural. Generating each speaker separately and stitching the takes together loses exactly that.<\/p>\n<\/div>\n                <\/div>\n                <div class=\"faq_divider\"><\/div>\n            <\/div>\n                    <div class=\"faq_item \">\n                <button type=\"button\" class=\"faq_question\" aria-expanded=\"false\">\n                    <span class=\"faq_question_text\">Can I control pauses, emphasis and pronunciation?<\/span>\n                    <svg class=\"faq_chevron\" width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                        <path d=\"M6 9L12 15L18 9\" stroke=\"currentColor\" stroke-width=\"1.5\" stroke-linecap=\"round\" stroke-linejoin=\"round\"\/>\n                    <\/svg>\n                <\/button>\n                <div class=\"faq_answer\" aria-hidden=\"true\" data-collapsed>\n                    <div class=\"faq_answer_content\"><p>Yes. Punctuation sets the pauses, capital letters add emphasis, and respelling a word phonetically fixes a stubborn pronunciation. On Eleven v3, audio tags give you direct control over delivery on top of all three.<\/p>\n<\/div>\n                <\/div>\n                <div class=\"faq_divider\"><\/div>\n            <\/div>\n                    <div class=\"faq_item \">\n                <button type=\"button\" class=\"faq_question\" aria-expanded=\"false\">\n                    <span class=\"faq_question_text\">Why does the voice change partway through a long recording?<\/span>\n                    <svg class=\"faq_chevron\" width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                        <path d=\"M6 9L12 15L18 9\" stroke=\"currentColor\" stroke-width=\"1.5\" stroke-linecap=\"round\" stroke-linejoin=\"round\"\/>\n                    <\/svg>\n                <\/button>\n                <div class=\"faq_answer\" aria-hidden=\"true\" data-collapsed>\n                    <div class=\"faq_answer_content\"><p>Usually the script was split across generations with a different model or voice on one part, or split mid-sentence so a fragment had no context to set its tone. Keep every part identical and break at paragraph ends. For long-form work, Multilingual v2 is the steadier of the two.<\/p>\n<\/div>\n                <\/div>\n                <div class=\"faq_divider\"><\/div>\n            <\/div>\n            <\/div>\n<\/section>\n\n<script type=\"application\/ld+json\">\n{\n    \"@context\": \"https:\/\/schema.org\",\n    \"@type\": \"FAQPage\",\n    \"mainEntity\": [\n        {\n            \"@type\": \"Question\",\n            \"name\": \"Which ElevenLabs model should I use for text to speech?\",\n            \"acceptedAnswer\": {\n                \"@type\": \"Answer\",\n                \"text\": \"Eleven v3 for storytelling, character work and anything that needs a performed read, since it is the model that responds to audio tags. Eleven Multilingual v2 for voiceovers, audiobooks and long recordings that need a consistent voice. Both cost the same in Picsart, so the choice is about the read rather than the budget.\"\n            }\n        },\n        {\n            \"@type\": \"Question\",\n            \"name\": \"How do audio tags work in ElevenLabs?\",\n            \"acceptedAnswer\": {\n                \"@type\": \"Answer\",\n                \"text\": \"Write them inline in your script inside square brackets, such as [whispers], [laughs] or [slowly], placed exactly where the delivery should change. Eleven v3 handles the full range of emotion, direction and audio effects. Multilingual v2 handles basic pauses and breaks only.\"\n            }\n        },\n        {\n            \"@type\": \"Question\",\n            \"name\": \"What is the ElevenLabs character limit?\",\n            \"acceptedAnswer\": {\n                \"@type\": \"Answer\",\n                \"text\": \"Eleven v3 takes 5,000 characters per generation, roughly five minutes of speech. Eleven Multilingual v2 takes 10,000, roughly ten minutes. For longer content, split the script across several generations.\"\n            }\n        },\n        {\n            \"@type\": \"Question\",\n            \"name\": \"Can ElevenLabs generate a conversation between two speakers?\",\n            \"acceptedAnswer\": {\n                \"@type\": \"Answer\",\n                \"text\": \"Yes. Eleven v3 supports natural multi-speaker dialogue, where the speakers share context inside a single generation so they react to one another and the timing sounds natural. Generating each speaker separately and stitching the takes together loses exactly that.\"\n            }\n        },\n        {\n            \"@type\": \"Question\",\n            \"name\": \"Can I control pauses, emphasis and pronunciation?\",\n            \"acceptedAnswer\": {\n                \"@type\": \"Answer\",\n                \"text\": \"Yes. Punctuation sets the pauses, capital letters add emphasis, and respelling a word phonetically fixes a stubborn pronunciation. On Eleven v3, audio tags give you direct control over delivery on top of all three.\"\n            }\n        },\n        {\n            \"@type\": \"Question\",\n            \"name\": \"Why does the voice change partway through a long recording?\",\n            \"acceptedAnswer\": {\n                \"@type\": \"Answer\",\n                \"text\": \"Usually the script was split across generations with a different model or voice on one part, or split mid-sentence so a fragment had no context to set its tone. Keep every part identical and break at paragraph ends. For long-form work, Multilingual v2 is the steadier of the two.\"\n            }\n        }\n    ]\n}<\/script>\n\n<script>\n(function() {\n    var container = document.getElementById('faq-faq-6a6c146c784e6');\n    if (!container) return;\n\n    var items = container.querySelectorAll('.faq_item');\n    items.forEach(function(item) {\n        var button = item.querySelector('.faq_question');\n        var answer = item.querySelector('.faq_answer');\n        if (!button || !answer) return;\n\n        button.addEventListener('click', function() {\n            var isActive = item.classList.contains('faq_item--active');\n\n            if (isActive) {\n                item.classList.remove('faq_item--active');\n                button.setAttribute('aria-expanded', 'false');\n                answer.setAttribute('aria-hidden', 'true');\n                answer.setAttribute('data-collapsed', '');\n            } else {\n                items.forEach(function(other) {\n                    var otherBtn = other.querySelector('.faq_question');\n                    var otherAnswer = other.querySelector('.faq_answer');\n                    other.classList.remove('faq_item--active');\n                    if (otherBtn) otherBtn.setAttribute('aria-expanded', 'false');\n                    if (otherAnswer) {\n                        otherAnswer.setAttribute('aria-hidden', 'true');\n                        otherAnswer.setAttribute('data-collapsed', '');\n                    }\n                });\n                item.classList.add('faq_item--active');\n                button.setAttribute('aria-expanded', 'true');\n                answer.removeAttribute('data-collapsed');\n                answer.setAttribute('aria-hidden', 'false');\n            }\n        });\n    });\n})();\n<\/script>\n\n<h2><span id=\"Start_generating_voiceovers\">Start generating voiceovers<\/span><\/h2>\n<p>ElevenLabs text to speech earns its place on the everyday work: narration for an explainer, a voiceover for a reel, a two-person script for a product walkthrough. Open the <a href=\"https:\/\/picsart.com\/ai-playground\/\">Picsart AI Playground<\/a>, pick the model that matches the read, and write your script so the punctuation does half the directing for you.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>ElevenLabs text to speech turns a written script into studio-quality spoken audio, and the model you pick decides what kind of performance you get. Eleven v3 acts a script with emotion and direction. Eleven Multilingual v2 narrates it cleanly and holds steady across long recordings. Both cost the same, so the choice is about the &hellip; <\/p>\n<p class=\"link-more\"><a href=\"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/\" class=\"more-link\">Continue reading<span class=\"screen-reader-text\"> &#8220;How to use ElevenLabs text to speech in Picsart&#8221;<\/span><\/a><\/p>\n","protected":false},"author":146,"featured_media":244536,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"_yoast_wpseo_title":"How to use ElevenLabs text to speech","_yoast_wpseo_metadesc":"Learn how to use ElevenLabs text to speech in Picsart: which model to pick, how audio tags direct the delivery, and how to generate a voiceover in a few steps.","faq_show":true,"faq_enable_schema":true,"how_to_show":true,"how_to_show_on_single":false,"how_to_enable_schema":true,"how_to_is_upload":true,"faq_title":"Get answers to common questions","how_to_title":"How to use ElevenLabs text to speech with Picsart","how_to_layout":"default","how_to_cta_text":"Try it in the AI Playground","how_to_cta_url":"https:\/\/picsart.com\/ai-playground\/","how_to_image_alt":"","how_to_display_image":0,"faq_items":null,"how_to_steps":null,"prompt_box_show":false,"prompt_box_placeholder":"","prompt_box_deeplink":"","prompt_box_submit_label":"","footnotes":""},"categories":[3181,1669],"tags":[],"class_list":["post-260798","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai","category-inspiration","entry"],"acf":{"footer_banner_name":"Start your design in Picsart","footer_banner_link_":"\/","footer_banner_button_text_":"Get Started","faq_show":true,"faq_title":"Get answers to common questions","faq_enable_schema":true,"faq_items":[{"question":"Which ElevenLabs model should I use for text to speech?","answer":"Eleven v3 for storytelling, character work and anything that needs a performed read, since it is the model that responds to audio tags. Eleven Multilingual v2 for voiceovers, audiobooks and long recordings that need a consistent voice. Both cost the same in Picsart, so the choice is about the read rather than the budget."},{"question":"How do audio tags work in ElevenLabs?","answer":"Write them inline in your script inside square brackets, such as [whispers], [laughs] or [slowly], placed exactly where the delivery should change. Eleven v3 handles the full range of emotion, direction and audio effects. Multilingual v2 handles basic pauses and breaks only."},{"question":"What is the ElevenLabs character limit?","answer":"Eleven v3 takes 5,000 characters per generation, roughly five minutes of speech. Eleven Multilingual v2 takes 10,000, roughly ten minutes. For longer content, split the script across several generations."},{"question":"Can ElevenLabs generate a conversation between two speakers?","answer":"Yes. Eleven v3 supports natural multi-speaker dialogue, where the speakers share context inside a single generation so they react to one another and the timing sounds natural. Generating each speaker separately and stitching the takes together loses exactly that."},{"question":"Can I control pauses, emphasis and pronunciation?","answer":"Yes. Punctuation sets the pauses, capital letters add emphasis, and respelling a word phonetically fixes a stubborn pronunciation. On Eleven v3, audio tags give you direct control over delivery on top of all three."},{"question":"Why does the voice change partway through a long recording?","answer":"Usually the script was split across generations with a different model or voice on one part, or split mid-sentence so a fragment had no context to set its tone. Keep every part identical and break at paragraph ends. For long-form work, Multilingual v2 is the steadier of the two."}],"how_to_show":true,"how_to_show_on_single":false,"how_to_title":"How to use ElevenLabs text to speech with Picsart","how_to_layout":"default","how_to_steps":[{"step_title":"Open the AI Playground","step_description":"Go to the Picsart AI Playground and switch to Audio mode. Everything runs in the browser, so there is nothing to install and nothing to set up.","show_cta_button":true},{"step_title":"Choose your model","step_description":"Pick Eleven v3 for an expressive, performed read, or Eleven Multilingual v2 for steady narration and a longer script in one pass. Auto Mode will otherwise choose a model for you.","show_cta_button":false},{"step_title":"Paste your script","step_description":"Add the words you want spoken. Punctuate deliberately, because commas and full stops create the pauses, and spell out acronyms and numbers the way you want them said. On Eleven v3 you can write audio tags inline in square brackets.","show_cta_button":false},{"step_title":"Set the language and accent","step_description":"Both are plain text fields rather than menus, so describe what you need. Naming a regional accent here is what produces one.","show_cta_button":false},{"step_title":"Pick a voice","step_description":"Audition two or three voices on a couple of sentences of your real script rather than on a sample line. A voice that fumbles a product name will keep fumbling it, and that only shows up on your own copy.","show_cta_button":false},{"step_title":"Generate and save","step_description":"Select Generate to see the credit cost, then save the audio to a project board so you can reuse it across designs and videos.","show_cta_button":false}],"how_to_enable_schema":true,"how_to_is_upload":true,"how_to_cta_text":"Try it in the AI Playground","how_to_cta_url":"https:\/\/picsart.com\/ai-playground\/","how_to_display_image":"","how_to_image_alt":"","prompt_box_show":false,"prompt_box_placeholder":"","prompt_box_deeplink":"https:\/\/picsart.com\/create\/editor?category=miniapps&app=com.picsart.aura","prompt_box_submit_label":"Create","try_prompt_show":false,"try_prompt_title":"Try this prompt","try_prompt_text":"","try_prompt_deeplink":"","tips_show":false,"tips_title":"Tips for best results","tips_items":null,"cta_banner_show":false,"cta_banner_title":"Need more space?","cta_banner_subtitle":"Extend any image in any direction with AI.","cta_banner_button_label":"Expand image","cta_banner_button_url":"","related_tools_title":"Related tools","related_tools_items":null,"post_level":""},"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v25.5 - https:\/\/yoast.com\/wordpress\/plugins\/seo\/ -->\n<title>How to use ElevenLabs text to speech<\/title>\n<meta name=\"description\" content=\"Learn how to use ElevenLabs text to speech in Picsart: which model to pick, how audio tags direct the delivery, and how to generate a voiceover in a few steps.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"How to use ElevenLabs text to speech\" \/>\n<meta property=\"og:description\" content=\"Learn how to use ElevenLabs text to speech in Picsart: which model to pick, how audio tags direct the delivery, and how to generate a voiceover in a few steps.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/\" \/>\n<meta property=\"og:site_name\" content=\"Picsart Blog\" \/>\n<meta property=\"article:publisher\" content=\"https:\/\/www.facebook.com\/picsart\" \/>\n<meta property=\"article:published_time\" content=\"2026-07-30T23:45:00+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-07-30T23:51:20+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/cdnblog.picsart.com\/2026\/03\/CR6776.-How-to-Make-an-AI-Voice-with-Picsart-_-1200_800.png\" \/>\n\t<meta property=\"og:image:width\" content=\"1200\" \/>\n\t<meta property=\"og:image:height\" content=\"800\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/png\" \/>\n<meta name=\"author\" content=\"Julia Tovmasyan\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:creator\" content=\"@PicsArtStudio\" \/>\n<meta name=\"twitter:site\" content=\"@PicsArtStudio\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Julia Tovmasyan\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"5 minutes\" \/>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"How to use ElevenLabs text to speech","description":"Learn how to use ElevenLabs text to speech in Picsart: which model to pick, how audio tags direct the delivery, and how to generate a voiceover in a few steps.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/","og_locale":"en_US","og_type":"article","og_title":"How to use ElevenLabs text to speech","og_description":"Learn how to use ElevenLabs text to speech in Picsart: which model to pick, how audio tags direct the delivery, and how to generate a voiceover in a few steps.","og_url":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/","og_site_name":"Picsart Blog","article_publisher":"https:\/\/www.facebook.com\/picsart","article_published_time":"2026-07-30T23:45:00+00:00","article_modified_time":"2026-07-30T23:51:20+00:00","og_image":[{"width":1200,"height":800,"url":"https:\/\/cdnblog.picsart.com\/2026\/03\/CR6776.-How-to-Make-an-AI-Voice-with-Picsart-_-1200_800.png","type":"image\/png"}],"author":"Julia Tovmasyan","twitter_card":"summary_large_image","twitter_creator":"@PicsArtStudio","twitter_site":"@PicsArtStudio","twitter_misc":{"Written by":"Julia Tovmasyan","Est. reading time":"5 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/#article","isPartOf":{"@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/"},"author":{"name":"Julia Tovmasyan","@id":"https:\/\/picsart.com\/blog\/ko\/#\/schema\/person\/74b70f3125250c23596a5306775b702d"},"headline":"How to use ElevenLabs text to speech in Picsart","datePublished":"2026-07-30T23:45:00+00:00","dateModified":"2026-07-30T23:51:20+00:00","mainEntityOfPage":{"@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/"},"wordCount":945,"publisher":{"@id":"https:\/\/picsart.com\/blog\/ko\/#organization"},"image":{"@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/#primaryimage"},"thumbnailUrl":"https:\/\/cdnblog.picsart.com\/2026\/03\/CR6776.-How-to-Make-an-AI-Voice-with-Picsart-_-1200_800.png","articleSection":["AI","Inspirational"],"inLanguage":"en-US"},{"@type":"WebPage","@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/","url":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/","name":"How to use ElevenLabs text to speech","isPartOf":{"@id":"https:\/\/picsart.com\/blog\/ko\/#website"},"primaryImageOfPage":{"@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/#primaryimage"},"image":{"@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/#primaryimage"},"thumbnailUrl":"https:\/\/cdnblog.picsart.com\/2026\/03\/CR6776.-How-to-Make-an-AI-Voice-with-Picsart-_-1200_800.png","datePublished":"2026-07-30T23:45:00+00:00","dateModified":"2026-07-30T23:51:20+00:00","description":"Learn how to use ElevenLabs text to speech in Picsart: which model to pick, how audio tags direct the delivery, and how to generate a voiceover in a few steps.","breadcrumb":{"@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/#primaryimage","url":"https:\/\/cdnblog.picsart.com\/2026\/03\/CR6776.-How-to-Make-an-AI-Voice-with-Picsart-_-1200_800.png","contentUrl":"https:\/\/cdnblog.picsart.com\/2026\/03\/CR6776.-How-to-Make-an-AI-Voice-with-Picsart-_-1200_800.png","width":1200,"height":800,"caption":"ai voice picsart"},{"@type":"BreadcrumbList","@id":"https:\/\/picsart.com\/blog\/how-to-use-elevenlabs-text-to-speech\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/picsart.com\/blog\/"},{"@type":"ListItem","position":2,"name":"How to use ElevenLabs text to speech in Picsart"}]},{"@type":"WebSite","@id":"https:\/\/picsart.com\/blog\/ko\/#website","url":"https:\/\/picsart.com\/blog\/ko\/","name":"Picsart Blog","description":"Keep up with the latest news in photo editing, digital photography, and art trends.","publisher":{"@id":"https:\/\/picsart.com\/blog\/ko\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/picsart.com\/blog\/ko\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/picsart.com\/blog\/ko\/#organization","name":"PicsArt Inc.","url":"https:\/\/picsart.com\/blog\/ko\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/picsart.com\/blog\/ko\/#\/schema\/logo\/image\/","url":"https:\/\/cdnblog.picsart.com\/2016\/02\/PicsArt-logo.png","contentUrl":"https:\/\/cdnblog.picsart.com\/2016\/02\/PicsArt-logo.png","width":195,"height":43,"caption":"PicsArt Inc."},"image":{"@id":"https:\/\/picsart.com\/blog\/ko\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/www.facebook.com\/picsart","https:\/\/x.com\/PicsArtStudio","https:\/\/www.instagram.com\/picsart","https:\/\/www.linkedin.com\/company\/picsart-photo-studio","https:\/\/www.pinterest.com\/picsart"]},{"@type":"Person","@id":"https:\/\/picsart.com\/blog\/ko\/#\/schema\/person\/74b70f3125250c23596a5306775b702d","name":"Julia Tovmasyan","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/picsart.com\/blog\/ko\/#\/schema\/person\/image\/","url":"https:\/\/cdnblog.picsart.com\/2026\/03\/3285C16C-FD87-4868-A2F0-04B6A0815CE1-150x150.jpg","contentUrl":"https:\/\/cdnblog.picsart.com\/2026\/03\/3285C16C-FD87-4868-A2F0-04B6A0815CE1-150x150.jpg","caption":"Julia Tovmasyan"}}]}},"featured_image":{"url":"https:\/\/cdnblog.picsart.com\/2026\/03\/CR6776.-How-to-Make-an-AI-Voice-with-Picsart-_-1200_800.png","dimensions":[]},"_links":{"self":[{"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/posts\/260798","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/users\/146"}],"replies":[{"embeddable":true,"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/comments?post=260798"}],"version-history":[{"count":10,"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/posts\/260798\/revisions"}],"predecessor-version":[{"id":260808,"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/posts\/260798\/revisions\/260808"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/media\/244536"}],"wp:attachment":[{"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/media?parent=260798"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/categories?post=260798"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/picsart.com\/blog\/wp-json\/wp\/v2\/tags?post=260798"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}