A story worth hearing. A video worth finishing. A customer who needs an answer, not another menu. ElevenLabs connects expressive AI audio, creative production and conversational agents. Start with the thing you want to make possible.
Affiliate links. FreeMalta may earn a commission at no additional cost to you. Features, usage and rights depend on the product and plan.
MAKE SOMETHING PEOPLE CAN FEEL. BUILD SOMETHING PEOPLE CAN USE.
01 / Two doors. Different jobs.
Are you making a story or starting a conversation?
A voiceover delivers your message. An agent listens and responds. Both can sound natural; only one needs permission to look up a booking or update a record. Choose the workflow before choosing the voice.
01
I want to create.
For creators, educators, publishers and marketing teams. Narrate a tutorial, build a campaign, localise a product video or give a character a voice.
The first useful win: one finished, reviewed asset. Not twenty unfinished generations.
For businesses and builders. Connect a voice or chat agent to approved knowledge and systems, with a clear route to a person when it reaches its limits.
The first useful win: one narrow task completed correctly. Not a convincing voice guessing its way through your business.
The brief might be a lesson, an advert, a short film or a product launch. ElevenCreative brings audio and visual tools into one workspace. You still direct the result: the argument, the taste, the timing and the final cut.
Voice / Text to speech
Give the words a delivery.
A training script needs clarity. A character needs personality. Text-to-speech turns your writing into narration; voice design, voice changing and authorised cloning provide different ways to shape the performance.
Try first: 30 seconds containing a name, a number, an acronym and a difficult sentence. Listen for meaning, not just realism.
Generate visuals with supported image and video models, then bring the assets into Studio. Add narration, music and sound effects to turn separate outputs into a coherent piece.
Know the distinction: the platform combines ElevenLabs audio with third-party visual models. Availability, costs and terms differ by model.
A strong explanation should not stop at the first language it was recorded in. Dubbing can help adapt spoken content for another audience, with a human reviewer checking meaning and delivery.
Try first: one short product tutorial in one additional language. Check names, local phrasing and timing. Do not assume dubbing includes lip sync.
Explore music generation for a scene, a sequence or a sonic identity. Sound effects serve another job: place the audience in the room, not merely fill the silence.
Before release: check the specific music licence. Advertising, film, games and other distribution uses can require additional rights; a subscription alone is not a universal clearance.
Picture a small team launching a service. One clear script. A short visual sequence. A recognisable voice. A music bed that knows when to get out of the way.
Studio is where those decisions come together. Review pacing, captions, pronunciation and rights before exporting. No generation button can decide what deserves to stay.
Illustration only. No audio or video is generated on this page.
03 / The interaction room
A good voice gets attention. A useful agent earns trust.
ElevenAgents lets you configure conversation, knowledge and workflows, then connect to business tools. Think of it as the interaction layer around a carefully designed job—not a substitute for knowing how that job works.
Where it can make sense
Repetition with a clear answer.
A front desk answering approved questions and routing requests.
A support team looking up information through permitted tools.
A service business qualifying an enquiry before a person takes over.
An internal help flow guiding staff to approved information.
These are possible workflows, not promises that every integration is available out of the box.
Knowledge: what is approved, current and in scope?
Actions: what can it read, change or trigger?
Fallback: when does it stop and reach a person?
Testing: how does it handle ambiguity and tool failures?
Data: who can access recordings and transcripts?
Tell people they are interacting with AI. Decide what should never be automated before opening the microphone.
The goal is not to sound human. It is to be useful without pretending.
Start small: one task, a test environment, approved data and human escalation. Measure correct outcomes, mistakes and handovers—not only how pleasant the conversation sounds. FreeMalta's implementation approach is editorial guidance, not a vendor performance guarantee.
04 / Why this company is worth watching
From a voice model to a broader interaction platform.
The interesting story is the expansion: speech became a creative workflow, and a voice became a way to interact with business systems. Growth is context—not a reason to skip your own product test.
2022 / The beginning
Voice first.
Founded in 2022. The starting point was AI speech; the catalogue has since expanded well beyond narration.
28 September 2026
v4 + Turbo
New speech models focused on expressive delivery and a lower-latency conversational variant. They are not video models.
30 September 2026
$22 billion
Valuation in a $300 million employee tender offer, following an $11 billion Series D valuation in February. Not $22 billion raised.
Milestones are based on ElevenLabs announcements reviewed 5 October 2026. Company valuation is neither a quality score nor a customer return.
05 / Make the first experiment useful
One afternoon. One thing worth keeping.
You do not need to adopt the whole platform. Pick a task you can finish, review and judge against the way you do it today.
Name the result.
“A clear 45-second product explanation” is a brief. “Try AI video” is an activity. Define the audience, the outcome and the person who will approve it.
Test the difficult bit.
For creative work, include names and awkward phrasing. For an agent, include missing information and a failed tool call. The easiest demo tells you very little.
Check what you can ship.
Review the actual plan, model access, export options, usage costs and rights. A demo that works technically may still be unsuitable for commercial release.
Keep it—or walk away.
Compare the reviewed result, time spent and total cost with your existing process. Adopt it if it improves the real workflow. Novelty alone is not a business case.
06 / The part that should not be hidden
Your voice is personal. Your licence is specific.
Permission is not a checkbox to rush.
Use recordings you are entitled to upload. Voice ownership, consent and product verification matter. Professional Voice Cloning is restricted to your own verified voice; another person's consent does not let you create their PVC yourself.
Do not use a familiar voice to imply endorsement, hide impersonation or make a misleading message more believable.
Paid does not mean unrestricted.
The general free plan is not licensed for commercial use. Paid plans remain subject to terms, model-specific restrictions and the rights in your inputs.
Music can need additional licensing for advertising and other uses. Check third-party visual model terms too. For customer data, establish recording consent where required, retention rules and access controls before deployment.
When another route may be better: use a human performer when the production needs bespoke artistic direction or rights arrangements. Keep sensitive or high-stakes conversations with qualified people where safe automation and escalation cannot be established.
07 / Budget for the workflow, not the headline
What will the finished result cost?
Include iterations, not just the successful generation. Creative credits, model choices, agent usage, telephony, integrations and licence needs can affect your total. Check current pricing for the product you are actually deploying.
For creative work
Estimate how much narration, dubbing, music and visual generation you need. Add time for revisions and human review. Check plan-specific commercial rights and any extra licences.
For agents
Estimate conversation volume, average duration, model and connection costs. Include setup, monitoring and escalation. “Cheaper per minute” is not useful if the task is completed incorrectly.
These FreeMalta guides describe complementary workflows, not claims of native integrations or bundled products.
09 / Before you press generate
A few good questions.
Is ElevenLabs only a text-to-speech tool?
No. Its offering spans speech generation, voice tools, music, sound effects, dubbing, transcription, creative image and video workflows, and conversational agents. The useful starting point is the job you need done, not the complete product catalogue.
What are Eleven v4 and v4 Turbo?
They are text-to-speech models announced on 28 September 2026. Eleven v4 focuses on expressive delivery; v4 Turbo is a lower-latency variant suited to conversational use. They are speech models, not the names of the video-generation models. Choose a model after testing your own script, language and pronunciation.
Does ElevenLabs make its own video models?
The creative workspace combines ElevenLabs audio technology with image and video models from other providers. Model availability, credit usage, regional access and output terms can differ. Do not assume that every visual model in the platform is developed by ElevenLabs.
Can I use the free plan for commercial content?
The general free plan does not include a commercial licence. Paid access does not mean unrestricted use of every output: check the relevant model, plan and use case. Music, third-party visual models, advertising and wider distribution can have additional conditions. Confirm rights before publication, not after.
Can I clone someone else’s voice?
Voice cloning requires the appropriate rights and permissions. Professional Voice Cloning is stricter: ElevenLabs says you can only create a Professional Voice Clone of your own voice, even if someone else consents. A voice owner can verify their own clone and share it through supported controls. Never treat permission as a reason to bypass verification.
Will an agent know my business automatically?
No. You need to configure its knowledge, instructions, allowed actions, integrations and escalation path. Test incorrect requests, missing information and tool failures. A natural voice is not evidence that the agent gave the right answer or completed the right action.
Can dubbing replace a native-language reviewer?
No. Use a reviewer to check meaning, names, local phrasing, subtitles and timing. Dubbing and lip sync are separate capabilities; lip sync is not automatically included in the dubbing workflow. Start with one language and one short clip.
Which subscription should I buy?
Choose around your actual output, model, commercial rights and usage. Creative generation and agent deployment have different cost drivers. Check current credits, limits, overages, add-ons and renewal terms on the relevant pricing screen. This page does not promise a fixed price or a promotional allowance.
The next move
Make it heard. Make it worth hearing.
Choose one story to tell or one conversation to improve. Give it a clear brief, a real test and a human standard. That is where an impressive tool becomes a useful part of your work.
FreeMalta may earn a commission through ElevenLabs links at no additional cost to you. This is a researched product guide, not a hands-on benchmark or a guarantee of output quality, compliance, integration availability or cost. Features and terms can change.
Research notes / reviewed 5 October 2026
ElevenLabs: ElevenCreative; ElevenAgents; Studio; current pricing and billing documentation.
Research announcement: Introducing Eleven v4, 28 September 2026, updated 1 October 2026.
Company announcements: Series D, 4 February 2026; employee tender at $22 billion valuation, 30 September 2026.
Help Centre: commercial use of generated content; Professional Voice Cloning restrictions; dubbing and lip-sync distinctions.
Product links on this page use FreeMalta's supplied affiliate destinations. Workflow illustrations and pilot suggestions are FreeMalta editorial guidance. No cloned voices, generated audio or unlicensed vendor demos are embedded.