APIPod logoAPIPod

Official full-capability at 5% off · Seedance 2.5

Seedance 2.5: 30-second storytelling, audio and picture generated together

ByteDance's next-generation audio-video joint generation model — built for 30-second narratives with precise reference control.

Seedance 2.5 doubles the single-clip ceiling to 4–30 seconds and raises the multimodal reference budget to 30 images, 10 videos and 10 audios per request. On APIPod you call the official full-capability model at 5% off list price through one unified API — discount and special-offer routes share the same endpoint, and plain public URLs are reviewed for you automatically.

Max duration

30s per clip

4–30s per generation; extend for longer cuts

Native audio

Audio + video

Dialogue, ambience and music synced to the frame

Reference budget

30 / 10 / 10

Up to 30 images, 10 videos, 10 audios per request

API routes

6 models

t2v / i2v / r2v × discount / special-offer

Seedance 2.5 vs 2.0 — what changed?

Same Seedance DNA; twice the runtime and a much larger reference budget.

Move to Seedance 2.5 when…

  • The story needs more than 15 seconds — vlogs, ads, dialogue scenes and one-take action up to 30s
  • Consistency demands a full reference pack: character sheets, locations, motion clips and voice tracks in one request
  • You want native audio — spoken lines, ambience and music generated with the picture
  • You need the official full-capability model at the lowest per-token cost available

Stay on Seedance 2.0 when…

  • Final delivery needs 4K ultra HD — the 2.5 discount tier tops out at 1080P; use the 2.0 standard tier for 4K (now live on APIPod)
  • You rely on the fast / mini / lite tiers for specialized speed or cost profiles
  • Existing 2.0 prompts already produce the shots you need at 5–15s

Spec comparison at a glance

Seedance 2.5 extends the 2.0 contract instead of replacing it — the model field is the only change most integrations need.

SpecSeedance 2.5Seedance 2.0
Duration per clip4–30s (default 4s)4–15s (default 5s)
Resolution tiersDiscount: 480P / 720P / 1080P; special-offer: 720P onlyUp to 4K
Reference imagesUp to 30Up to 9
Reference videosUp to 10Up to 3
Reference audiosUp to 10Up to 3
Native audioYes — on by defaultYes

Both generations ship through APIPod's full-capability pipeline: media URLs are reviewed by the asset service before submission, so plain public URLs work in requests. Billing is token-based on successful tasks only — see the pricing page for live rates.

Official list price · 5% off

Official pricing reference

Seedance 2.5 bills by tokens on the official side: token consumption per second scales with resolution, and a video reference is billed for the reference duration together with the output duration at a lower unit price.

ResolutionTokens per secondWithout video referenceWith video reference
480P9,607.5 tokens/s$10.7 /M tokens · ≈ $0.103/s$6.4 /M tokens · ≈ $0.062/s
720P21,600 tokens/s$10.7 /M tokens · ≈ $0.231/s$6.4 /M tokens · ≈ $0.138/s
1080P≈ 38,500 tokens/s$10.7 /M tokens · ≈ $0.412/s$6.4 /M tokens · ≈ $0.246/s

With-video billing counts reference-video duration plus output duration as tokens. Official example — 10s reference + 10s output: 480P $1.23, 720P $2.76 per clip.

List prices from the official Seedance 2.5 pricing page; APIPod's discount tier bills at 5% off. The special-offer routes (seedance-2.5-lite-*) output 720P only — 1080P is not available there. Your actual rates on APIPod are shown on the pricing page.

What Seedance 2.5 is built for

Longer narratives, better control — the official positioning, delivered through one API.

30-second one-take stories

Write the clip as timed beats — setup, action, payoff — and Seedance 2.5 holds blocking, wardrobe and scene logic across the full take. Extend twice for even longer cuts.

Reference-driven consistency

Up to 30 images, 10 videos and 10 audios per request: character identity, product appearance, motion paths and voice tracks each get an explicit job in the prompt.

Audio-video joint generation

Dialogue, ambience and music are generated together with the visuals and land on the frame — including invented-language performances and perfectly timed SFX.

Professional camera language

Orbits, whip pans, speed ramps, handheld tracking and invisible-cut staging: name the move and the model executes it with cinematic framing.

Recommended workflow: draft wide, finalize once

1) Draft at 480P

Validate composition, timing and reference roles cheaply — and test a 5–10s opening before committing to the full 30s.

2) Borrow proven prompts

Start from the community templates in the prompt library; each detail page explains the beat structure so you can rebuild the rhythm for your subject.

3) Finalize at 720P / 1080P

Rerun the winning prompt at 720P or 1080P with the same reference pack; fix a single failed beat by editing only that time range instead of regenerating.

4) One API, every Seedance

Switch between 2.5 and 2.0 tiers by changing the model field — auth, billing and observability carry over.

Example request

Seedance 2.5 quick start

Public model IDs on one endpoint: discount routes seedance-2.5-t2v for text-to-video, seedance-2.5-i2v for first/last-frame animation and seedance-2.5-r2v for multimodal references; special-offer seedance-2.5-lite-* routes bill at special-offer rates (720P only).

Media URLs are reviewed by the asset service automatically — pass plain public URLs and the pipeline exchanges them for approved asset references. Tasks are asynchronous; poll the task ID until completed.

curl -X POST https://api.apipod.ai/v1/videos/generations \
  -H "Authorization: Bearer $APIPOD_API_KEY" \
  -H "Content-Type: application/json" \
  -d {
    "model": "seedance-2.5-t2v",
    "prompt": "A cinematic tracking shot through a rain-soaked neon street, realistic motion, synchronized ambient sound.",
    "duration": 8,
    "resolution": "720p",
    "aspect_ratio": "16:9",
    "generate_audio": true
  }

Seedance 2.5 FAQ

What is Seedance 2.5?

Seedance 2.5 is ByteDance's next-generation audio-video joint generation model, built for 30-second storytelling. It generates video with synchronized native audio, follows timed-beat prompts, and accepts large multimodal reference packs for identity, motion and voice control.

How long can a Seedance 2.5 video be?

A single generation runs 4–30 seconds (default 4s). For longer narratives, write the clip as timed beats and extend the result — the official guidance allows extending twice for longer cuts.

What is the difference between Seedance 2.5 and 2.0?

2.5 doubles the per-clip ceiling from 15s to 30s and raises reference limits from 9 images / 3 videos / 3 audios to 30 / 10 / 10. On resolution, the discount routes output 480P / 720P / 1080P and the special-offer routes (seedance-2.5-lite-*) are locked to 720P, while Seedance 2.0 supports up to 4K. Prompt structure, native audio and the asset-review pipeline are shared.

Does Seedance 2.5 support 1080P or 4K?

Yes for 1080P. The discount routes (seedance-2.5-t2v / i2v / r2v) output 480P / 720P / 1080P; the special-offer routes (seedance-2.5-lite-*) are locked to 720P. For 4K masters, use the Seedance 2.0 standard tier (4K now live on APIPod).

Does Seedance 2.5 generate audio?

Yes. Audio-video joint generation is on by default (generate_audio=true): dialogue, ambience and music are produced together with the visuals. You can disable it per request, and steer voices and sound with reference audio clips on the r2v route.

How do I call Seedance 2.5 through the API?

POST to /v1/videos/generations with model set to seedance-2.5-t2v, seedance-2.5-i2v or seedance-2.5-r2v. The task is asynchronous — poll the returned task ID until status is completed. Media URLs you pass are reviewed by the asset service automatically, so plain public URLs just work.

Tell 30-second stories with Seedance 2.5

Official full-capability model at 5% off, deeper special-offer pricing, one unified API — your first clip is one request away.