三十秒是我们提供的最长单次生成时长,仅与 Wan 3.0:https://openrouter.ai/alibaba/wan-3.0 和 Wan 3.0 Prime:https://openrouter.ai/alibaba/wan-3.0-prime 相当。其他所有公开持续时间范围的模型最长不超过20秒。这之所以重要,是因为较长的视频替代方案是拼接短片,而拼接容易导致连续性中断。如果你的交付物是30秒的广告或单个连续场景,这条路线是最短的路径。
这些请求按公式计费,而不是固定的每秒费率,这是你在扩展前最需要了解的一点。代币数量是(宽度 x 高度 x fps x 时长)/ 1024,24帧每秒,每个代币收费为0.00000107美元,或者当请求带有视频参考时为0.0000064美元。所以唯一的变量是像素、秒数,以及你是否提供素材,你可以在播放前给任何片段定价。
Seedance 2.5 has been live on our video API since August 7, 2026, and as of September 3, 2026, the model page:https://openrouter.ai/bytedance/seedance-2.5 lists it from $0.1028 per second of generated video, which is what 480p works out to. That per-second figure is derived, not fixed. Billing is per video token, and the token count scales with output pixels as well as duration, so a second of 720p costs a little over twice a second of 480p from the same model.
The model is also a different shape from Seedance 2.0. It runs to 30 seconds instead of 15, and it stops at 720p instead of 4K. Below is what it does well, what a clip costs at each resolution, how it compares to Seedance 2.0, Wan 3.0, and Veo 3.1 on our own catalog data, and the cases where we’d recommend a different model.
Seedance 2.5 is ByteDance’s video generation model, and it takes several kinds of input into one video output. You can generate from a text prompt alone, from images that fix the first or last frame of a shot, or from reference assets that steer the result without pinning a specific frame. Our model page describes it as suited to long-form storytelling, reference-based generation, video editing, and video extension, and those four jobs cover most of what it’s used for.
Here’s the current spec snapshot from our video models endpoint:https://openrouter.ai/api/v1/videos/models, so you can see where those capabilities stop in practice.
Last checked September 3, 2026 on the video models endpoint.
The endpoint also lists the exact output sizes, which is useful when you’d rather send size than a resolution and an aspect ratio: 854x480, 752x560, 640x640, 560x752, 480x854, 992x432 at 480p, and 1280x720, 1112x834, 960x960, 834x1112, 720x1280, 1470x630 at 720p.
The model name is the human-readable label, ByteDance: Seedance 2.5, while the slug is the string you put in the model field, bytedance/seedance-2.5 . We carry five, and they aren’t interchangeable. Seedance 2.5 is the newest and the longest. Seedance 2.0:https://openrouter.ai/bytedance/seedance-2.0 is the higher-resolution model, with clips of 4 to 15 seconds at 480p through 4K for $0.000007 per video token at 480p and 720p. Seedance 2.0 Fast:https://openrouter.ai/bytedance/seedance-2.0-fast and Seedance 2.0 Mini:https://openrouter.ai/bytedance/seedance-2.0-mini run to 15 seconds at a 720p ceiling for $0.0000042 and $0.0000035 per token, which makes them the draft slugs. Seedance 1.5 Pro:https://openrouter.ai/bytedance/seedance-1-5-pro is the older model that generates video and audio in a single pass, stops at 1080p and 12 seconds, and bills $0.0000024 per token with audio or half that without.
Seedance 2.5 lists no 1080p and no 4K output. If you need either, that’s Seedance 2.0 in the same family, or Veo 3.1 outside it. This is the one spec where the newer version is narrower than the older one, and it’s worth checking against your delivery format before you build around 2.5.
The 4-to-30 second window and the reference inputs describe what the model accepts. The reason to choose it over the rest of the family is narrower than that. It’s the Seedance model for length and for footage you already have.
Four optional controls, each reducing how much the model can change on its own. Frame images and reference assets select different modes rather than combining, and frame images take priority if both are sent.
Thirty seconds is the longest single generation we carry, matched only by Wan 3.0:https://openrouter.ai/alibaba/wan-3.0 and Wan 3.0 Prime:https://openrouter.ai/alibaba/wan-3.0-prime. Every other model that publishes a duration range stops at 20 seconds or less. That matters because the alternative to a long take is stitching short ones, and stitching is where continuity breaks down. If your deliverable is a 30-second spot or a single continuous scene, this is the shortest path to it.
input_references :https://openrouter.ai/docs/cookbook/video-generation/reference-to-video accepts image, audio, and video assets on Seedance generation 2 and newer, which includes 2.5. Image references carry a face, a product, or a style. A video reference gives the model existing footage to edit or extend. An audio reference gives it a track to work against. Our model page lists up to 50 reference assets per request, which is a model page figure rather than a limit our endpoint publishes, so treat the exact ceiling as reported rather than measured.
Video and audio references work on every Seedance model from generation 2 onward, so this isn’t unique to 2.5 inside the family. It does separate the family from the rest of the catalog. Wan 3.0, Veo 3.1, and Kling v3.0 Pro all take image references only.
A request that carries a video reference and no frame images bills at $0.0000064 per video token instead of $0.0000107, about 40% less. At 720p that’s $0.138 per second instead of $0.231. So the cheapest way to get 30 seconds of 720p out of this model is to extend footage you already have rather than to generate it cold.
frame_images :https://openrouter.ai/docs/cookbook/video-generation/image-to-video accepts entries typed as first_frame and last_frame , so you can fix where a shot begins and ends and let the model fill in the motion between them. Wan 3.0, the other 30-second model, lists first_frame only. If a request carries both frame images and reference assets, the frame images take priority and the job is treated as image-to-video, so the references will have no visible effect and the request bills at the base rate.
generate_audio defaults to true, and we price video tokens identically whether audio is on or off, so muting a generation saves nothing. Veo 3.1 charges $0.40 per second with audio against $0.20 without, and Seedance 1.5 Pro halves its own rate for silent output. The model page also lists multilingual audiovisual generation, so a line in a language other than English is worth attempting here before you budget for a separate voice pass.
Length, references, and included audio fit a specific kind of job: one continuous scene, or a change to footage that already exists. Three cases fit that description well.
This is the strongest fit, since 30 seconds is enough for a full scene and the model doesn’t need you to cut. Write the beats in order and give the camera one instruction per beat.
Send the clip as a video reference and describe only what should change or what comes next. This is the request shape that bills at the lower video-input rate.
Short lines sync better than long ones, but 30 seconds gives you room for two or three of them in one take instead of one line per generation.
Those prompts go into a request body rather than a chat message, because video generation doesn’t run on /chat/completions . It has a dedicated asynchronous endpoint:https://openrouter.ai/blog/tutorials/video-generation-api, so you submit a job to POST /api/v1/videos , poll the polling_url we return until the status reads completed , then download the result with your API key. Generation usually takes from 30 seconds to a few minutes, and a 30-second polling interval is a reasonable default. Video models also don’t appear in the plain models list, so use /api/v1/videos/models or the video model collection:https://openrouter.ai/collections/video-models to find them.
The minimal text-to-video call is a submit and a poll.
Adding references turns that into an editing or extension workflow. The example below extends an existing clip, which is the request shape that bills at the lower video-input rate. Resolution, duration, and aspect ratio are all optional, and the seed value is a placeholder rather than a required value.
Seedance 2.5 accepts a seed, but determinism isn’t guaranteed by every provider. Check the video generation docs:https://openrouter.ai/docs/guides/overview/multimodal/video-generation before you build a workflow that depends on repeatable output. Provider-specific options:https://openrouter.ai/docs/cookbook/video-generation/provider-specific-video-options travel under provider.options. .parameters , where is the provider slug. Seedance runs on one provider, seed , and its allowed passthrough keys are watermark , req_key , and output_format .
Those requests bill on a formula rather than a flat per-second rate, which is the single most useful thing to understand before you scale up. The token count is (width x height x fps x duration) / 1024 at 24 fps, and every token bills at $0.0000107, or $0.0000064 when the request carries a video reference. So the only variables are pixels, seconds, and whether you’re supplying footage, and you can price any clip before you run it.
Last checked September 3, 2026. Figures are calculated from the token formula and the per-token rates on the video models endpoint. Vertical formats cost the same as landscape at equal pixel counts, so 720x1280 prices match 1280x720.
Working one row through end to end, a 10-second 720p clip in 16:9 is (1280 x 720 x 24 x 10) / 1024, which comes to 216,000 video tokens, and 216,000 tokens at $0.0000107 is $2.31. The per-second column is that figure divided by duration and rounded to three decimals.
Two details are worth knowing before you reconcile a bill. Our up-front estimate uses the standard dimensions for each resolution tier, so a 21:9 clip at 720p is estimated as if it were 1280x720, and the final charge comes from the token count the provider reports when the job completes. A request that carries a video reference is authorized at a flat $2 instead of the formula, because the token count depends on the length of your input footage, which we don’t know until the job runs. The final charge for those requests still comes from the reported token count at the video-input rate.
Measured against Veo 3.1 with audio at $0.40 per second, Seedance 2.5 at 480p is about four times cheaper and at 720p about 1.7 times cheaper. Measured against Wan 3.0 at $0.05 per second at 480p and $0.10 at 720p, Seedance 2.5 costs about twice as much at both. It isn’t the cheapest way to generate video. It’s the cheapest way to generate a 30-second clip with last-frame control, video references, or audio references.
Once you can price a clip, the comparison comes down to which requirement you hit first:https://openrouter.ai/docs/cookbook/video-generation/choose-video-model. All four models below are in our catalog, and every figure comes from the same endpoint on the same day.
The five categories creators test, against what our own sources document for each one.
Last checked September 3, 2026 on the video models endpoint. Seedance per-second figures are derived from the token formula, while Wan 3.0 and Veo 3.1 publish flat per-second SKUs on that same endpoint, so the figures are comparable but reached by different routes.
Seedance 2.5 wins on inputs and on length together. It’s the only model here that takes 30 seconds, both frame types, and video and audio references in the same request. Seedance 2.0 is the one to use for 1080p and 4K, and it’s cheaper per second at every resolution the two share. Wan 3.0 matches the 30-second ceiling at half the price and reaches 1080p, but it lists first-frame control only and no video or audio references, so it’s the better default for long text-to-video and the worse one for editing. Veo 3.1 is the fixed-length option, at $0.60 per second for 4K, and Veo 3.1 Fast generates 4K for $0.30 per second if you want the cheaper route to that resolution.
Because all four sit behind the same POST /api/v1/videos call and the same key, running your own comparison is a one-field change rather than four subscriptions. The video model collection:https://openrouter.ai/collections/video-models lists everything we carry with the modality filter already applied.
You don’t have to write code to compare them. Our video benchmarks page:https://openrouter.ai/benchmarks/media/videos runs the same prompt through every video model we carry, including Seedance 2.5, Seedance 2.0, Wan 3.0, and Veo 3.1, and shows each output next to what it cost and how long it took to generate. Watch the clips, sort by cost or generation time, then pick the models you want to try in chat. Use it to check the claims in this review before you spend anything.
Seedance 2.5 is the right choice when clip length, existing footage, or included audio matter more than resolution. Thirty seconds in one generation, image, video, and audio references, both frame types, and a lower rate for video-referenced requests are a combination nothing else in our catalog offers. Use Seedance 2.0 when you need 1080p or 4K, Wan 3.0 when you want 30 seconds of text-to-video for less, and Veo 3.1 when a fixed 4, 6, or 8-second 4K clip is the deliverable. Teams that need repeatable, approval-ready output should treat its results as material to review rather than as finals.
Yes, for longer single takes and for work that starts from existing footage. Seedance 2.5 generates clips of 4 to 30 seconds, accepts image, video, and audio references, supports first- and last-frame control, and generates audio in the same pass at the same rate. It’s a weaker choice when you need 1080p or 4K output, the lowest price per second, or frame-exact reproducibility.
Seedance 2.5 generates video from a text prompt, from images that fix the first or last frame, or from reference assets that carry a subject, a style, or existing footage into a new clip. As of September 3, 2026, it produces clips of 4 to 30 seconds at 480p or 720p, across six aspect ratios, with audio by default and a seed parameter to narrow variance between runs. The model page:https://openrouter.ai/bytedance/seedance-2.5 also lists video editing and video extension, and up to 50 reference assets per request.
Seedance 2.5 bills per video token, and tokens are (width x height x fps x duration) / 1024 at 24 fps, so the price scales with both pixels and seconds. The rate is $0.0000107 per token, which works out to about $0.103 per second at 480p and $0.231 at 720p, so a 10-second 720p clip costs roughly $2.31. Audio is included at the same rate. A request that carries a video reference bills at $0.0000064 per token instead, about 40% less.
There’s no single best video model, and the right answer depends on which requirement you hit first. Seedance 2.5 is the pick for 30-second clips that need last-frame control or video and audio references. Seedance 2.0 is the pick for 1080p and 4K. Wan 3.0 also runs to 30 seconds, costs less per second, and reaches 1080p, but it lists first-frame control only. Veo 3.1 is the pick for fixed 4, 6, or 8-second clips at 4K. All of them run through the same asynchronous video endpoint, so you can compare them on your own prompt by changing one field. The video model collection:https://openrouter.ai/collections/video-models lists the full set.
Send a POST request to https://openrouter.ai/api/v1/videos with bytedance/seedance-2.5 in the model field and your prompt in prompt , then poll the polling_url from the response until the status is completed and download the result with your API key. Generation is asynchronous and usually takes from 30 seconds to a few minutes. The full request schema, including webhooks:https://openrouter.ai/docs/cookbook/video-generation/video-generation-webhooks, is on the video generation page:https://openrouter.ai/docs/guides/overview/multimodal/video-generation.
Yes. Seedance 2.5 generates audio in the same pass as the video, and generate_audio defaults to true. We bill the same per-token rate with audio on or off, so there’s no cost reason to disable it. That isn’t true of Veo 3.1 or Seedance 1.5 Pro, which both charge less for silent output. The model page also lists multilingual audiovisual generation, so it’s worth testing a line in your own target language.
Seedance 2.5 runs twice as long and stops at a lower resolution. It generates 4 to 30 second clips at 480p or 720p, accepts image, video, and audio references, and bills $0.0000107 per video token. Seedance 2.0:https://openrouter.ai/bytedance/seedance-2.0 generates 4 to 15 second clips at 480p through 4K and bills $0.000007 per token at 480p and 720p, so it’s cheaper per second and the only one of the two that reaches 1080p or 4K. Pick 2.5 for length and for editing existing footage, and 2.0 for resolution.
All product sources above were last checked on September 3, 2026. Creator coverage is cited as framing only, not as a source of specs or prices.