Seed Audio 1.0
Sign up
ByteDance next audio model · early-bird window

Seed Audio 2.0: dialogue, music, and effects in one pass.

Seed Audio 2.0 is built for T2A, TA2A, TV2A, and TAV2A. One brief can hold six minutes, six reference voices, thirty languages, and independent stems for dialogue, music, ambience, and effects. Reserve credits now. The studio below still generates with Seed Audio 1.0, so you can practice prompts before Seed Audio 2.0 goes live.

T2ATA2ATV2ATAV2A
6 min
longest scene window at launch
6 refs
reference voices the new model can keep
30 langs
language coverage in the published spec

Scene desk preview

Sample brief

Coming soon
Dark teal control room preview for Seed Audio 2.0 with waveform monitors

A captain, engineer, and officer argue as alarms climb, engines rumble, and strings tighten before impact.

Dialogue stem
Music stem
Ambience + FX

Rehearse Seed Audio 2.0 prompts in the live studio.

Seed Audio 2.0 appears in the model menu with a Coming Soon tag. Generation still runs on Seed Audio 1.0, so you can test credits, references, and scene briefs before Seed Audio 2.0 opens. Keep the same account; early-bird orders apply when the new model is switched on.

Model

Input

Prompt-first audio generation with optional controls.

Prompt*
83/2048

Additional Settings

Customize your input with more control.

History

Your recent Seed Audio 1.0 Preview generations.

Sign in to see your generation history.

How to use

How to brief Seed Audio 2.0 while the 1.0 studio is live.

Treat today’s generator as a rehearsal room for Seed Audio 2.0. Write the scene, attach the references you already have, set output limits, then listen. When Seed Audio 2.0 ships, the same briefing habit should transfer: longer takes, more voices, video context, and stems instead of a single bounce.

  1. Prompt editor used to rehearse Seed Audio 2.0 scene briefs
    01

    Write the scene, not a slogan

    Name who speaks, where they stand, what the room does, and which cue must land. That is the brief the new model is being trained to follow at longer lengths.

  2. Reference audio and image controls in the current studio
    02

    Park the references you will reuse

    Today the studio accepts fewer beds than the 2.0 spec. Still upload the voices you care about so the prompt mentions them the same way you will after launch.

  3. Output settings for format, speed, and credit cap
    03

    Cap credits before you iterate

    Set format, speed, and a credit ceiling. Seed Audio 2.0 will ask for longer jobs; practicing cheap drafts now keeps the habit honest.

  4. Generated result with playback and download controls
    04

    Listen like an editor, not a fan

    Check overlap, music masking, and late Foley. Those are the notes stems and timestamps are meant to reduce.

Seed Audio 2.0 brief

Mode + duration + speakers + language + picture or voice refs + stem intent + timestamped cues.

Open the 1.0 studio

After launch

What you can ship with Seed Audio 2.0

Seed Audio 2.0 is not a longer TTS clip. It is a scene renderer: picture-aware dubbing, extra reference voices, and stems you can drop on a timeline. The four boards below are the jobs we hear producers asking for while they wait for Seed Audio 2.0.

Rainy alley film still for a Seed Audio 2.0 short-drama example

01

Multi-character short drama

Keep several speakers in one take, then leave rain, footsteps, and a sting on their own tracks so an editor can ride them independently.

Dubbing suite with a film monitor for Seed Audio 2.0 video-to-audio

02

Video-aware dubbing

Feed a picture plus optional voice beds. The 2.0 spec is to follow action, cuts, and pacing instead of laying a flat narration under silent video.

Fantasy forest still used to illustrate Seed Audio 2.0 game audio

03

Playable world beds

Boss lines, forest loops, and UI hits can share one brief, then split so a game mixer does not have to unbake a stereo bounce.

Night-time car commercial still for a Seed Audio 2.0 advertising example

04

Campaign spots with stems

A three-voice car spot can keep VO, engine, city wash, and the hook on separate stems for localization and legal recuts.

Where it lands

Seed Audio 2.0 use cases

These are production jobs, not mood boards. Seed Audio 2.0 is being positioned for people who already cut picture, localize spots, or ship interactive worlds and need the soundtrack to arrive as editable parts. If you only need a single narrator, Seed Audio 1.0 or a speech engine may still be simpler until Seed Audio 2.0 is live.

01

Films and comic dramas

Aimed at multi-scene shorts where voices, beds, and hits have to stay in character across cuts instead of being restitched from stock.

02

Ads and brand films

Use it when a spot needs VO, product Foley, and a music lift in one pass, then separate stems for market versions.

03

Games and interactive worlds

Barks, loops, and stingers can be briefed together. Stems give audio leads something closer to a session than a trailer bounce.

04

Localization and dubbing

Thirty-language coverage plus video context is the Seed Audio 2.0 pitch for teams that currently re-record every market by hand.

05

Longer narrative audio

Six minutes is still a chapter, not a series, but it is enough for a podcast cold open or a comic-drama act that Seed Audio 1.0 could not hold in one job.

06

Crowded ensemble scenes

Six reference voices help keep a cast distinct instead of collapsing everyone into two stock timbres.

Model notes

Seed Audio 2.0 characteristics

The list below is the published Seed Audio 2.0 delta versus the 1.0 studio you can run today: more time, more references, picture input, stems, tighter cues, and a wider language set. It is a production spec, not a slogan sheet.

Four input modes

Specified for T2A, TA2A, TV2A, and TAV2A, so a job can start from text, voice beds, picture, or all three.

Six reference audios

The bed count rises from about three to six, which matters when a scene has a lead, a foil, and a crowd texture.

Picture as context

TV2A and TAV2A read action and pacing instead of guessing a score under a silent timeline.

Independent stems

Dialogue, music, ambience, and effects can leave Seed Audio 2.0 as separate tracks instead of one glued mix.

Timestamped cues

Seed Audio 2.0 is described with precise cue placement so a door slam or a line pickup can sit on a marked beat.

Thirty languages

Wider language coverage for dubbed spots and localized drama without resetting the entire mix.

Version delta

How Seed Audio 2.0 differs from Seed Audio 1.0

Seed Audio 1.0 is what you can generate here today. Seed Audio 2.0 is the upcoming Dreamina-class model: longer jobs, more references, video input, stems, and a wider language set. Use the table as a briefing sheet, not a promise that every 2.0 control is already in this studio.

Feature comparison between Seed Audio 1.0 and Seed Audio 2.0
DimensionSeed Audio 1.0 (July 2026)Seed Audio 2.0 (current Dreamina spec)
Maximum durationAbout 2 minutesUp to 6 minutes
Reference audiosUp to about 3 clipsUp to 6 clips
Input methodsText + reference audio (+ optional image)Text + reference audio + reference video (T2A / TA2A / TV2A / TAV2A)
Language support20+ languages30 languages
Multi-track and time controlBasic timeline control (dialogue precision about 100ms)Independent stems for dialogue, music, ambience, and effects, plus precise timestamps
Core capabilitiesEnd-to-end scene generation (dialogue + SFX + ambience + music)Stronger controllable cloning, emotion/rhythm/style control, video-aware dubbing, and cross-scene voice consistency on top of the 1.0 scene model
Best suited forShort scenes and basic sound sketchingLonger audio, short dramas, ads, games, multilingual localization, and crowded multi-character scenes

Seed Audio 2.0 is not a free upgrade inside this generator yet. Order early-bird credits if you want capacity waiting when the model is enabled; keep using Seed Audio 1.0 for drafts.

Lock early-bird pricing

Pricing

Choose the right Seed Audio 1.0 plan

Subscribe for the best value, or buy credits when you need a flexible top-up. Every paid plan and credit pack includes Seed Audio 1.0 API access, with one shared credit balance across the web app and API.

Free
$0

Try Seed Audio 1.0 with free signup credits. Perfect for testing prompts and short scenes.

  • 10 free credits
  • Up to ~8 seconds per generation
  • Max 2 min audio per generation
  • API access
  • Priority support
Pro
Save $39.98
$16.66$19.99/mo

The best annual choice for creators with ongoing audio needs.

  • 30,000 credits per year
  • Up to ~400 total audio minutes
  • Everything in Free
  • More room for longer audio-scene projects
  • Reference voice uploads
  • Priority support
  • Max 2 min audio per generation
  • Seed Audio 1.0 API access — shared credits
Max
Save $99.98
$41.66$49.99/mo

The strongest value for teams that expect heavy audio generation.

  • 90,000 credits per year
  • Up to ~1,200 total audio minutes
  • Everything in Pro
  • Best value for high-volume generation
  • Larger annual production buffer
  • Team and agency-friendly usage
  • Max 2 min audio per generation
  • Seed Audio 1.0 API access — shared credits
Questions about billing, access, or custom needs?contact@seedaudio2.co

Field notes

What people actually say while waiting on Seed Audio 2.0

These are composite notes from trailer editors, localization producers, and game audio leads — written to sound like working notes, not a five-star wall. Nobody here claims Seed Audio 2.0 has shipped on this site.

I bought credits because six-minute stems would save a weekend of temp dubs. Until Seed Audio 2.0 is on, I still run 1.0 just to see whether the prompt is even sane.
Mara K. · Freelance trailer editor
The video-to-audio pitch is the only reason I care. If the new model can follow a cut instead of ignoring it, we stop laying English VO under picture that already has mouth flaps.
Jun Park · Localization producer, Seoul
I do not need another magic voice. I need Seed Audio 2.0 to keep the heroine on one bed across three episodes. If cloning still drifts, we will keep recording.
Elena V. · Comic-drama audio lead
Six references is the upgrade I can explain to a producer. Three was always a squeeze. I will believe the stems when I can solo the rain without killing the line.
Ravi S. · Game audio contractor

Early-bird

Hold a Seed Audio 2.0 seat before the price moves.

Seed Audio 2.0 is listed as coming soon. Orders placed now keep early-bird pricing. You can still generate with Seed Audio 1.0 on this page while the new model is queued.