APIPod logoAPIPod

Official full-capability · Seedance 2.5

Seedance 2.5: 30-second storytelling, audio and picture generated together

ByteDance's next-generation audio-video joint generation model — built for 30-second narratives with precise reference control.

Seedance 2.5 doubles the single-clip ceiling to 4–30 seconds and raises the multimodal reference budget to 30 images, 10 videos and 10 audios per request. On APIPod you call the official full-capability model through one unified API — discount and special-offer routes share the same endpoint, and plain public URLs are reviewed for you automatically.

Max duration

30s per clip

4–30s per generation; extend for longer cuts

Native audio

Audio + video

Dialogue, ambience and music synced to the frame

Reference budget

30 / 10 / 10

Up to 30 images, 10 videos, 10 audios per request

API routes

6 models

t2v / i2v / r2v × discount / special-offer

Seedance 2.5 vs 2.0 — what changed?

Same Seedance DNA; twice the runtime and a much larger reference budget.

Move to Seedance 2.5 when…

  • The story needs more than 15 seconds — vlogs, ads, dialogue scenes and one-take action up to 30s
  • Consistency demands a full reference pack: character sheets, locations, motion clips and voice tracks in one request
  • You want native audio — spoken lines, ambience and music generated with the picture
  • You need the official full-capability model at the lowest per-token cost available

Stay on Seedance 2.0 when…

  • Final delivery needs 1080P+ — the 2.5 family tops out at 720P; the 2.0 standard tier ships 1080P today with 4K rolling out
  • You rely on the fast / mini / lite tiers for specialized speed or cost profiles
  • Existing 2.0 prompts already produce the shots you need at 5–15s

Spec comparison at a glance

Seedance 2.5 extends the 2.0 contract instead of replacing it — the model field is the only change most integrations need.

SpecSeedance 2.5Seedance 2.0
Duration per clip4–30s (default 4s)4–15s (default 5s)
Resolution tiersDiscount: 480P / 720P; special-offer: 720P onlyUp to 4K
Reference imagesUp to 30Up to 9
Reference videosUp to 10Up to 3
Reference audiosUp to 10Up to 3
Native audioYes — on by defaultYes

Both generations ship through APIPod's full-capability pipeline: media URLs are reviewed by the asset service before submission, so plain public URLs work in requests. Billing is token-based on successful tasks only — see the pricing page for live rates.

Official list price

Official pricing reference

Seedance 2.5 bills by tokens on the official side: token consumption per second scales with resolution, and a video reference is billed for the reference duration together with the output duration at a lower unit price.

ResolutionTokens per secondWithout video referenceWith video reference
480P9,607.5 tokens/s$10.7 /M tokens · ≈ $0.103/s$6.4 /M tokens · ≈ $0.062/s
720P21,600 tokens/s$10.7 /M tokens · ≈ $0.231/s$6.4 /M tokens · ≈ $0.138/s

With-video billing counts reference-video duration plus output duration as tokens. Official example — 10s reference + 10s output: 480P $1.23, 720P $2.76 per clip.

List prices from the official Seedance 2.5 pricing page. The special-offer routes (seedance-2.5-lite-*) output 720P only — 1080P is not available on the 2.5 family. Your actual rates on APIPod are shown on the pricing page.

What Seedance 2.5 is built for

Longer narratives, better control — the official positioning, delivered through one API.

30-second one-take stories

Write the clip as timed beats — setup, action, payoff — and Seedance 2.5 holds blocking, wardrobe and scene logic across the full take. Extend twice for even longer cuts.

Reference-driven consistency

Up to 30 images, 10 videos and 10 audios per request: character identity, product appearance, motion paths and voice tracks each get an explicit job in the prompt.

Audio-video joint generation

Dialogue, ambience and music are generated together with the visuals and land on the frame — including invented-language performances and perfectly timed SFX.

Professional camera language

Orbits, whip pans, speed ramps, handheld tracking and invisible-cut staging: name the move and the model executes it with cinematic framing.

Recommended workflow: draft wide, finalize once

1) Draft at 480P

Validate composition, timing and reference roles cheaply — and test a 5–10s opening before committing to the full 30s.

2) Borrow proven prompts

Start from the community templates in the prompt library; each detail page explains the beat structure so you can rebuild the rhythm for your subject.

3) Finalize at 720P

Rerun the winning prompt at 720P with the same reference pack; fix a single failed beat by editing only that time range instead of regenerating.

4) One API, every Seedance

Switch between 2.5 and 2.0 tiers by changing the model field — auth, billing and observability carry over.

Example request

Seedance 2.5 quick start

Public model IDs on one endpoint: discount routes seedance-2.5-t2v for text-to-video, seedance-2.5-i2v for first/last-frame animation and seedance-2.5-r2v for multimodal references; special-offer seedance-2.5-lite-* routes bill at special-offer rates (720P only).

Media URLs are reviewed by the asset service automatically — pass plain public URLs and the pipeline exchanges them for approved asset references. Tasks are asynchronous; poll the task ID until completed.

curl -X POST https://api.apipod.ai/v1/videos/generations \
  -H "Authorization: Bearer $APIPOD_API_KEY" \
  -H "Content-Type: application/json" \
  -d {
    "model": "seedance-2.5-t2v",
    "prompt": "A cinematic tracking shot through a rain-soaked neon street, realistic motion, synchronized ambient sound.",
    "duration": 8,
    "resolution": "720p",
    "aspect_ratio": "16:9",
    "generate_audio": true
  }

Seedance 2.5 FAQ

What is Seedance 2.5?

Seedance 2.5 is ByteDance's next-generation audio-video joint generation model, built for 30-second storytelling. It generates video with synchronized native audio, follows timed-beat prompts, and accepts large multimodal reference packs for identity, motion and voice control.

How long can a Seedance 2.5 video be?

A single generation runs 4–30 seconds (default 4s). For longer narratives, write the clip as timed beats and extend the result — the official guidance allows extending twice for longer cuts.

What is the difference between Seedance 2.5 and 2.0?

2.5 doubles the per-clip ceiling from 15s to 30s and raises reference limits from 9 images / 3 videos / 3 audios to 30 / 10 / 10. On resolution, the discount routes output 480P / 720P and the special-offer routes (seedance-2.5-lite-*) are locked to 720P, while Seedance 2.0 supports up to 4K. Prompt structure, native audio and the asset-review pipeline are shared.

Does Seedance 2.5 support 1080P or 4K?

No. The discount routes output 480P and 720P, and the special-offer routes (seedance-2.5-lite-t2v / i2v / r2v) are locked to 720P. For higher-resolution masters, use the Seedance 2.0 standard tier (1080P today, 4K rolling out on APIPod).

Does Seedance 2.5 generate audio?

Yes. Audio-video joint generation is on by default (generate_audio=true): dialogue, ambience and music are produced together with the visuals. You can disable it per request, and steer voices and sound with reference audio clips on the r2v route.

How do I call Seedance 2.5 through the API?

POST to /v1/videos/generations with model set to seedance-2.5-t2v, seedance-2.5-i2v or seedance-2.5-r2v. The task is asynchronous — poll the returned task ID until status is completed. Media URLs you pass are reviewed by the asset service automatically, so plain public URLs just work.

Tell 30-second stories with Seedance 2.5

Official full-capability model, discount and special-offer pricing, one unified API — your first clip is one request away.