Projects
Browser-based AI audio and video editor for voiceovers, dubbing, music, sound effects, and captions with Eleven v4.
If your output is narrated — podcasts, audiobooks, voiceover-led video — Studio's text-based editing removes the single most painful step in the process: re-recording a bad take. Eleven v4 landing in September 2026 also moves expressiveness and cloning quality meaningfully. Budget attention, though: the shared credit pool is the real constraint, not the feature list.
Verified 48m ago · liveness 75/100 · cite: rightaichoice.com/tools/projects
- Podcasters who want to fix a fluffed line by editing text, clean noisy dialogue, and score a theme without leaving one editor
- Audiobook authors importing EPUB/PDF/TXT manuscripts and casting multiple voices across long-form narration
- Video creators adding AI voiceover, auto captions, and generated background music to MP4/MOV edits
- Localization teams producing multilingual dubs and subtitles from a single shared project timeline
- Post-production houses needing multicam editing, color grading, or finishing-grade video tools
- Editors who require offline or desktop software — Studio runs in the browser
- Creators publishing long music- or dub-heavy projects on tight credit budgets
We scan live Reddit threads, YouTube comments, X posts, G2 reviews and other communities — and hand you an honest verdict in under a minute.
- Honest verdict, not marketing
- Real pros & cons from real users
- Attributed quotes with receipts
3 free scans · no card needed
Skip Studio if you need offline or desktop editing, multicam and color-grading finishing tools, or if your month leans heavily on Eleven Music (900 credits/min) or Dubbing Studio without watermark (10,000 credits/min) and a shared credit pool won't stretch.
Credits are shared across every ElevenLabs product, so music at 900 credits/min and Dubbing Studio without watermark at 10,000 credits/min drain the same pool your voiceovers draw from.
Studio's ladder runs Free ($0), Starter ($6/mo), Creator ($22/mo, first month $11), Pro ($99/mo), Scale ($299/mo), Business ($990/mo), and custom Enterprise. Annual billing is two months free — effectively $5/mo Starter, $18.33/mo Creator, $82.50/mo Pro, $249.17/mo Scale, and $825/mo Business. Solo creators get real value from Starter and Creator; teams needing seats should budget for Scale at $299/mo, which is pricier than Descript's team tiers but bundles voice, music, and dubbing into one
In short
Projects — Browser-based AI audio and video editor for voiceovers, dubbing, music, sound effects, and captions with Eleven v4. Best for Podcasters who want to fix a fluffed line by editing text, clean noisy dialogue, and score a theme without leaving one editor, Audiobook authors importing EPUB/PDF/TXT manuscripts and casting multiple voices across long-form narration, Video creators adding AI voiceover, auto captions, and generated background music to MP4/MOV edits. Free to start; paid plans from $6/mo.
What's new in Projects
Checked 5 days agoAcross the latest 10 updates: 4 feature updates, 3 launches, 1 changelog entry and 2 news mentions.
ElevenLabs valuation increases to $22 billion fueled by enterprise demand for conversational agents
ElevenLabs' valuation rose to $22B, attributed to enterprise demand for conversational agents.
Speech to Text transcript editing
Batch and realtime transcription now accept a natural-language edit instruction up to 2,000 characters; responses return an edited_transcript.
Introducing Eleven v4, our most emotive model
Eleven v4 and v4 Turbo add emotional nuance, faster response times and stronger voice cloning to TTS, live in ElevenAgents, ElevenCreative and the API.
Eleven v4 and Eleven v4 Turbo available
Eleven v4 ships with higher-quality expressive speech and improved voice cloning across 90+ languages; v4 Turbo targets real-time use at ~100 ms median latency.
Image & Video
Changelog entry dated September 23, 2026 covering image and video capabilities.
ElevenAgents image and video support
Changelog entry adding image and video support to ElevenAgents, with API and SDK updates.
ElevenAgents Music
Music support added to ElevenAgents alongside SDK releases, new endpoints and schema changes.
Scribe v2 Medical
Scribe v2 Medical released, extending the transcription model line to medical use cases.
Universal Music Group and ElevenLabs announce multi-year strategic agreement
ElevenLabs signed a multi-year strategic agreement with Universal Music Group.
ElevenAgents Workspaces
Workspaces added to ElevenAgents, with SDK releases, new endpoints, updated endpoints and schema changes.
What people actually say about Projects — is it worth it?
We ran a structured research pass across product reviews, community discussions, and post-purchase forum threads to surface the patterns vendors won't publish themselves. Below: the recurring strengths, the hidden costs people mention most, and the cohort that consistently regrets adopting this tool.
46 mentions across 4 sources (Hacker News, Product Hunt, Lemmy, Tech Press) · researched Jul 3, 2026.
Average across the 4 sources that answered — each source counts once, not each post.
- +Over 10,000 voices with natural intonation and emotion.
- +Text-based audio editing is intuitive and fast for voiceovers.
- +AI music generation (Music v2) creates high-quality, diverse tracks.
- +Voice cloning works well with minimal training audio.
- +Built-in sound effects generator saves time finding assets.
- −Audio/video desync issues on longer projects.
- −Credit system is confusing and easy to deplete.
- −Customer support is slow and unresponsive.
- −Video editing capabilities are very limited.
- −Speech correction sometimes misaligns with waveform.
- • Credits are shared across all ElevenLabs tools, so heavy use in one area drains quickly
- • Music generation and sound effects consume additional credits beyond TTS
Viability Score
How well maintained and how widely used is Projects? Built from what the vendor actually publishes (docs, changelog, tutorials, integrations, pricing), whether the site is live, and how much real users discuss it. How we calculate this
Last calculated: October 2026
How we score →Key Features
- Text to Speech with a 17,000+ voice library and 32+ languages in Studio
- Eleven v4 and v4 Turbo models: expressive speech, stronger cloning, 90+ languages, ~100 ms median latency on v4 Turbo
- Speech Correction: edit the script and regenerate the same cloned voice instead of re-recording
- Voice Isolator: remove background noise and reverb from audio and video dialogue
- Studio Agent: AI co-editor that drafts scripts, selects voices, places sound effects, and arranges clips
- Eleven Music generation and automatic video scoring on the timeline
- Prompt-based sound effect generation for ambience and cinematic impact
- One-click captions with custom styles and multilingual subtitles
- Voice Changer for transforming existing recordings in the browser
- Speech to Text transcription with natural-language transcript editing (Scribe v2 family)
- Dubbing with automatic and Dubbing Studio workflows for multilingual releases
- Video support: import MP4/MOV, trim, merge, and sync audio to picture
- Document import from EPUB, PDF, TXT, HTML, or a starting URL for audiobook projects
- Multi-cast character casting: assign different voices to individual text fragments
- Public project URLs with time-stamped timeline comments for client review
About Projects
ElevenLabs Studio is a browser-based AI audio and video editor built for people who produce spoken-word content: podcasters, audiobook authors, video creators, localization teams, and AI filmmakers. You bring a script, a manuscript, or an MP4/MOV file and assemble everything on one timeline — AI voiceovers pulled from the Voice Library of 17,000+ voices, Eleven Music soundtracks, prompt-generated sound effects, captions, voice conversion, transcription, and dubbing. Nothing to install; projects live in the browser and can be shared through public URLs. The workflow that separates Studio from a generic NLE is Speech Correction. Fluff a line and you don't re-record — you edit the text and Studio regenerates it in the same cloned voice. Voice Isolator strips background noise and reverb from dialogue, one-click captions add custom styles plus multilingual subtitles, and document import (EPUB, PDF, TXT, HTML, or a starting URL) turns a manuscript into a structured audiobook project. Multi-cast casting assigns distinct voices to individual text fragments, so a full cast comes out of one text document. Studio Agent is the co-editor layer: describe what you want and it drafts the script, picks voices, places sound effects, and arranges clips, with manual takeover available at any point. Latest news extends the editing surface further — Eleven v4 and v4 Turbo (September 28, 2026) push expressiveness, stronger voice cloning, and 90+ language coverage, with v4 Turbo tuned for real-time work at roughly 100 ms median latency, and batch plus realtime transcription now accept natural-language edit instructions and return an edited transcript. Pricing is credit-based and shared across every ElevenLabs product, so speech, music, dubbing, and transcription draw from one pool — check per-product burn rates before committing. Everything in Studio is reachable programmatically through the ElevenAPI.
Behind the Verdict
We'd reach for Studio when the hard part of the job is the voice, not the picture. Podcasters cleaning dialogue, authors revising a chapter's narration by retyping a sentence, video creators who need a voiceover plus captions plus a scored bed on one timeline — that's the sweet spot, and Speech Correction is why. Edit the text, keep the clone, skip the booth session. Document import into EPUB/PDF/TXT/HTML also makes long-form manuscript work feel native rather than bolted on. Pick it when you want one subscription covering speech, music, SFX, captions, dubbing, and transcription, and when Speech to Text being able to take a natural-language edit instruction (2,000 characters, returns an edited_transcript) saves your team another tool. The Eleven v4 arrival in late September 2026 matters here: 90+ languages, sharper cloning, and v4 Turbo around 100 ms median latency for real-time use. Pass if you're cutting multicam, color grading, or doing finishing-grade video work. Studio does timeline trim, merge, and sync — not compositing or visual effects. Pass too if you need offline or desktop operation, since it runs in the browser, and if your team needs advanced VFX beyond trimming. Watch the credits. Everything draws from one monthly pool and the burn is uneven: Voice Changer and Voice Isolator run 1,000 credits per minute, Eleven Music 900 per minute, Speech to Text 330 per minute, and Dubbing Studio without watermark 10,000 per minute. On Pro's 600,000 credits, a watermark-free dub of a 60-minute episode eats the month. Music- and dub-heavy producers should model spend before subscribing. The closest alternative depends on what you're optimizing. Descript trades AI-native voice tooling for a stronger transcript-first video editor; Adobe Premiere Pro and DaVinci
Researching Projects? Get your full AI stack in 60 seconds.
Free, no signup — tell us your goal and get tools matched to your budget & existing stack.
Real-world workflow fit
Concrete scenarios for the personas Projects actually fits — and what changes day-one when you adopt it.
Import an EPUB manuscript, cast a distinct voice per character with multi-cast, then spot a misread line in chapter 12 and fix it by editing the text so Speech Correction regenerates the same cloned voice.
Outcome: A completed audiobook project with consistent narration and no re-recording sessions, shareable via a public project URL for editor or client review.
Upload the raw episode, run Voice Isolator to strip room reverb and background noise, fix flubbed sentences with Speech Correction, then generate an original theme track with Eleven Music and export at 192 kbps.
Outcome: A cleaned episode with an original soundtrack, produced in one browser session instead of a DAW plus a separate music subscription.
Import an MP4, add an AI voiceover synced on the timeline, generate sound effects from prompts, add one-click multilingual captions, then produce a dub into a second language from the same project.
Outcome: One source project that ships a captioned original plus a localized dub version, with credit spend visible before rendering.
Use Cases
- Turn a book manuscript into a professional audiobook with distinct voices per character.
- Clean up a podcast episode, fix flubbed lines, and generate an original theme track.
- Add AI voiceover, auto captions, and scored background music to an MP4 or MOV edit.
- Produce multilingual dubs and subtitles from a single source project.
- Prototype a film scene with AI voiceover, sound effects, and music on one timeline.
- Describe a video brief to Studio Agent and get a drafted script with voices and clips arranged.
Models Under the Hood
as of 2026-10-08
Limitations
- Studio is browser-based, so offline editing isn't possible, and video support stops at MP4/MOV import with trim, merge, and sync — no multicam, color grading, or compositing.
- Every product draws from one shared monthly credit pool, and the vendor's own pricing FAQ lists Eleven Music at 900 credits per minute and Dubbing Studio without watermark at 10,000 credits per minute, so a $99/mo Pro plan's 600,000 credits can be consumed by a single long dub.
- Free is capped at 3 Studio projects and 10,000 credits, and paid tiers cap monthly credits (30k Starter, 121k Creator, 600k Pro, 1.8M Scale, 6M Business).
- Voice cloning is plan-gated — Instant Voice Cloning starts at Starter, Professional Voice Cloning at Creator.
- Team collaboration and workspace seats start at Scale ($299/mo, 3 seats), and custom SSO, DPAs/SLAs, and HIPAA BAAs require Enterprise.
- Unused credits roll over for up to two months only while you stay on an active paid plan — downgrading or cancelling forfeits them, and Free has no rollover.
as of 2026-09-27
Verification history
We have re-verified Projects 9 times since . Each pass re-reads the vendor's own pages and re-checks every listed field against that evidence; passes where nothing had changed are marked as such.
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
Showing the 6 most recent of 9 verification passes.
Free to cite with attribution — this page re-verifies continuously.
12-month cost
Project the real annual outlay, including the implied monthly cost when only an annual tier is published.
Vendor list price only. Add-on usage, seat overages, and contract minimums are surfaced under Hidden costs & gotchas.
Plans compared
For each published Projects tier: who it actually fits, and what it adds vs. the previous tier. Cross-reference the cost calculator above for projected annual outlay.
Free
$0/mo
Ideal for
Someone evaluating Studio before spending anything — a single creator testing text-to-speech, Speech Correction, and captions on short personal projects.
What this tier adds
Starting tier: $0 for 10,000 credits and 3 Studio projects, personal use only, 128 kbps MP3 export, no credit rollover.
Starter
$6/mo
Ideal for
A solo podcaster or hobbyist creator who needs a commercial license and their first cloned voice without committing to a $22/mo tier.
What this tier adds
Adds a commercial license, Instant Voice Cloning, 20 Studio projects, music commercial use, Dubbing Studio, and Image & Video; credits rise from 10k to 30k.
Creator
$22/mo
Ideal for
A working audiobook author or regular video creator producing multi-character work who needs Professional Voice Cloning and headroom to buy extra credits.
What this tier adds
Adds Professional Voice Cloning, additional credits purchasable, and 192 kbps export; credits jump from 30k to 121k.
Pro
$99/mo
Ideal for
A full-time creator or small studio running frequent long-form narration who needs higher audio fidelity on API output.
What this tier adds
Adds 44.1 kHz PCM audio output via API and 192 kbps quality audio; credits rise from 121k to 600k.
Scale
$299/mo
Ideal for
A small team — think a three-person podcast network or localization shop — that needs shared seats and team collaboration in the browser.
What this tier adds
First tier with workspace seats (3) and team collaboration, plus 3 Professional Voice Clones; credits rise from 600k to 1.8M.
Business
$990/mo
Ideal for
A 10-person production team or agency running high-volume dubbing and voice work who needs low-latency TTS economics.
What this tier adds
Adds low-latency TTS as low as 5c/minute, 10 workspace seats, and 10 Professional Voice Clones; credits rise from 1.8M to 6M.
Enterprise
Custom
Ideal for
Regulated or high-scale organizations — healthcare, enterprise media, or anyone whose security review blocks a listed tier — that need custom terms and assurance.
What this tier adds
Custom credits and seats plus DPAs/SLAs, BAAs for HIPAA, custom SSO, elevated concurrency limits, fully managed dubbing with Productions, and priority support.
Where the pricing makes sense
The company stage and team size where Projects's pricing actually pencils out — and where peers do it cheaper.
Studio's ladder runs Free ($0), Starter ($6/mo), Creator ($22/mo, first month $11), Pro ($99/mo), Scale ($299/mo), Business ($990/mo), and custom Enterprise. Annual billing is two months free — effectively $5/mo Starter, $18.33/mo Creator, $82.50/mo Pro, $249.17/mo Scale, and $825/mo Business. Solo creators get real value from Starter and Creator; teams needing seats should budget for Scale at $299/mo, which is pricier than Descript's team tiers but bundles voice, music, and dubbing into one
Setup time & first value
How long it actually takes to get something useful out of Projects — broken out by persona, not the marketing-page minute.
Solo creators can go from signup to a first rendered voiceover in about 10–15 minutes using Free or Starter — paste a script, pick a voice, hit generate. Audiobook authors should budget an hour or two for document import, speaker assignment, and multi-cast voice casting. Teams adding Scale or Business tiers need extra time for workspace setup, seat assignment, and professional voice clone
Switching to or from Projects
How to bring data in from common predecessors and how to get it back out — written for the switcher, not the buyer.
- →From Descript: export your audio and transcript, then rebuild the project in Studio and use Speech Correction plus Eleven Music for the generative assets Descript doesn't provide.
- →From Adobe Podcast: move cleaned audio and scripts into Studio to add Eleven Music scoring, dubbing, and multi-cast voice casting on one timeline.
- →From a manual DAW workflow: export stems and a script, import the script as TXT or PDF, and let Studio Agent draft the arrangement before you refine manually.
- →From a book manuscript: import the EPUB or PDF directly and assign speakers per section rather than slicing chapters by hand.
- ↗To Descript: export your Studio audio and captions, then rebuild for conventional multitrack editing and non-generative video workflows.
- ↗To Adobe Premiere or DaVinci Resolve: export stems, voiceover WAVs, and caption files to finish with multicam, color grading, and effects tools Studio doesn't cover.
- ↗To a standalone DAW: pull WAV exports at 44.1 kHz (Pro tier and above) into your mix session for fine-grained mastering.
Resources & Guides
Tutorials & Learning
YouTube returned 6 videos for “Projects”, and we withheld 6: 6 could not be judged, because “Projects” is a single word that other videos use for other things. We are showing none, because we could not prove any of them are about Projects.
Official links
Tools that pair well with Projects
Common stack mates teams adopt alongside Projects, with the specific reason each pairing earns its keep.
Invideo AI
Agentic AI video editor that applies one instruction across every shot on a multitrack timeline.
MimicPC
Browser-based cloud that runs 20+ pre-installed open-source AI apps for image, video, and audio generation
Magnific AI
Magnific is an AI creative suite with 30+ image, video, audio and 3D tools sharing one credit pool and a node-based Spaces canvas.
Featured Head-to-Head Comparisons
Projects vs Landr Mastering
Choose LANDR Mastering if you need pro-level AI mastering with album workflow and DAW integration; choose Projects (ElevenLabs) if you want an all-in-one AI audio/video editor for voiceovers, music, and sound effects. LANDR is Mastering-specific; Projects is broader but browser-only. For mastering only, LANDR wins on price and features.
Projects vs Splice
Choose Splice if you're a music producer who needs a massive royalty-free sample library and rent-to-own plugins like Serum 2. Choose Projects if you're a video/audio creator who wants to generate voiceovers, music, and sound effects from text. They serve different workflows; Splice is for sample-based music production, Projects for AI-assisted audio/video editing.
Projects vs Storyfile
Buy StoryFile if your goal is an authentic, interactive human experience using real footage for museums, legacies, or digital twins — its conversational AI is unrivaled for emotional and historical accuracy. Choose Projects (ElevenLabs Studio) if you need a fast, browser-based audio/video editor with AI voiceovers, music, and sound effects for content creation, where synthetic voices and text-based editing speed up production. They serve fundamentally different needs and are not direct competitors.
Alternatives to Projects
View allInvideo AI
Agentic AI video editor that applies one instruction across every shot on a multitrack timeline.
MimicPC
Browser-based cloud that runs 20+ pre-installed open-source AI apps for image, video, and audio generation
Magnific AI
Magnific is an AI creative suite with 30+ image, video, audio and 3D tools sharing one credit pool and a node-based Spaces canvas.
Frequently Asked Questions
Best-of guides
Used Projects? Help shape our editorial sentiment research.