Try MiniMax FastH3 V2 from Fastvideo →
Models
Agents
Workflows
Studio
PricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
Generative Media AgentCreate and edit media by chattingWorkflow AgentBuild visual workflows with Agent
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
Explore models/
Audio & SpeechActive

Eleven V4 by ElevenLabs (Text-to-Speech)

byelevenlabs

Eleven V4 by ElevenLabs turns scripts into emotionally rich speech with strong speaker identity and multilingual coverage. Add audio tags and export MP3 audio.

Text to Speech
Model ID
eleven v4
Provider
elevenlabs
Updated
1790690894
eleven v4
0
Comments
Average rating : 5 (2 users)
Providerelevenlabs
Modeleleven v4
Text to Speech
wiro playground—elevenlabs/eleven v4
Reset to defaults

Required. Up to 10,000 characters. Supports audio tags like [whispering], [laughing], [shouting], [excited].

Required.

Required.

Sample outputs
elevenlabs-eleven-v4-sample-1.mp3
elevenlabs-eleven-v4-sample-2.mp3
elevenlabs-eleven-v4-sample-3.mp3
elevenlabs-eleven-v4-sample-4.mp3
Updated 1790690894

Overview

Eleven V4 is an audio model from ElevenLabs for expressive text-to-speech. It reads your script, infers intent from context, then performs it like a voice actor. You pick a voice, and the model keeps that voice identity stable across long passages. This is useful when you need natural narration or character dialogue without recording sessions.

What you can build

  • Audiobook and long-form narration with emotional range
  • Character voiceovers for games, animation, and interactive stories
  • Multilingual voiceovers that keep the same speaker identity across languages
  • Podcast-style segments, intros, and ad reads from a written script
  • Dialogue scenes that include reactions like laughter, sighs, or whispering

Inputs

  • A script you want spoken, as plain text, up to 10,000 characters per request
  • Optional bracketed direction tags inside the script, like [whispering], [laughing], or [shouting]
  • A voice choice from the included preset voices
  • An MP3 export quality choice, based on sample rate and bitrate
  • An optional stability control from 0 to 1 to trade expressiveness for consistency
  • An optional similarity control from 0 to 1 to keep closer to the chosen voice

Outputs

The model returns a single spoken-audio file in MP3 format. The file contains the rendered performance of your script in the selected voice. If you include bracketed direction tags, the audio may include delivery changes and non-verbal reactions where they appear.

Limitations

  • One request supports up to 10,000 characters, which is roughly 10 minutes of audio
  • Output can vary between runs, even with the same script and settings
  • Audio tags are best-effort. Some tags may be ignored or applied inconsistently
  • Names, numbers, and abbreviations can be misread. Write the spoken form when it matters
  • Voice quality depends on the selected voice. Noisy or inconsistent cloned samples can add artifacts
  • Poorly structured text can reduce quality, including messy punctuation and copied PDF formatting

Safety & compliance

  • Only clone voices when you have the right to use that voice. Professional voice cloning requires verification
  • ElevenLabs provides detection tooling for ElevenLabs-generated audio, including watermark-based detection
  • Generated audio may still enable impersonation misuse. Use clear disclosure where required by policy or law
  • You retain ownership of generated audio, but commercial use depends on your ElevenLabs plan terms

Example prompts

Great starting points for eleven v4.

[measured] Every great story begins with a single voice. [softly] Tonight, let that voice be yours.Audio & Speech
[annoyed] You ate my sandwich again? [sighs] Fine. [shouting] But the coffee is off limits!Audio & Speech
[whispering] Did you hear that? [nervous] Something is moving in the kitchen. [laughing] Oh, it's just the cat.Audio & Speech
[tired] It's been a long day. [sighs] I think I'll just order pizza. [cheerful] Actually, make that two.Audio & Speech

API quick start

Run eleven v4 with a single API call.

POST https://api.wiro.ai/v1/Run/elevenlabs/eleven-v4
{
  "prompt": "[measured] Every great story begins with …",
  "voice": "george",
  "outputFormat": "mp3_44100_128",
  "stability": 0.5
}
curl
curl -X POST "https://api.wiro.ai/v1/Run/elevenlabs/eleven-v4" \
  -H "Content-Type: application/json" \
  -H "x-api-key: YOUR_WIRO_API_KEY" \
  --data-binary @- <<'JSON'
{
  "prompt": "[measured] Every great story begins with …",
  "voice": "george",
  "outputFormat": "mp3_44100_128",
  "stability": 0.5
}
JSON
View full API docs

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingPartnerChangelogStatusFAQ
ModelsNano Banana 2GPT Image 2.5Seedream V5 ProSeedance 2.5Veo 3.1Kling V3FLUX 3FLUX.2 ProWan 3.0 PrimeGrok Imagine 1.5
PartnersGoogleOpenAIByteDanceBlack Forest LabsKling AIQwenAlibabaxAIMiniMaxElevenLabs
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion