Seedance 2.5 Image to Video
Animate any image into a cinematic video up to 30 seconds with Seedance 2.5 by ByteDance. Start from a first-frame image, optionally set a last frame to control the ending, and get synchronized audio by default. Output preserves your image's aspect ratio.
42% less than the priciest provider (5 compared). Same model, same output.
First-frame image. The video starts from this image.
Optional last-frame image. The video transitions from the first frame to this one.
Duration
Duration of the video in seconds (4-30). Billed per second.
Resolution
Video resolution.
Audio
Synchronized audio: voice, sound effects and music generated in the same pass.
Content
What the video contains. Sellers differ in what they allow: 'general' (no faces) can unlock lower prices; 'mature' routes only to sellers without content restrictions.
42% less than the priciest provider (5 compared). Same model, same output.
Your result renders here.
Use Seedance 2.5 Image to Video via the API
Generate with Seedance 2.5 Image to Video from your own code. Async by design: submit a generation, get a run id back instantly, poll until it completes. Billed in credits at the best available provider price — from $0.14 per second, audio included.
1 — Start a generation
curl -X POST https://app.scenetra.com/api/v1/generate/seedance-2-5-i2v \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"prompt":"Slow push-in, the subject turns to camera and smiles, soft wind in the hair","image":"https://example.com/first-frame.jpg","duration":"5"}'Response (202)
{
"runId": "run_abc123",
"status": "PENDING",
"model": "seedance-2-5-i2v",
"estimatedCredits": 1350,
"statusUrl": "/api/v1/runs/run_abc123"
}2 — Poll until COMPLETED (every 2-5 seconds)
curl https://app.scenetra.com/api/v1/runs/run_abc123 \
-H "Authorization: Bearer YOUR_API_KEY"Response
{
"runId": "run_abc123",
"status": "COMPLETED",
"progress": { "completed": 1, "total": 1 },
"outputs": [
{ "type": "video", "url": "https://images.scenetra.com/..." }
],
"error": null,
"estimatedCredits": 1350
}Request fields
| Field | Type | Description | Default |
|---|---|---|---|
| prompt | string | Optional. Describe the desired motion and camera work; the look comes from the first frame. | — |
| image* | string (https URL) | First-frame image. The video starts from this image. | — |
| last_image | string (https URL) | Optional last-frame image. The video transitions from the first frame to this one. | — |
| duration | string | Duration of the video in seconds (4-30). Billed per second.4 · 5 · 6 · 7 · 8 · 9 · 10 · 11 · 12 · 13 · 14 · 15 · 16 · 17 · 18 · 19 · 20 · 21 · 22 · 23 · 24 · 25 · 26 · 27 · 28 · 29 · 30 | 5 |
| resolution | string | Video resolution.480p · 720p | 720p |
| generate_audio | string | Synchronized audio: voice, sound effects and music generated in the same pass.false · true | true |
| subject_content | string | What the video contains. Sellers differ in what they allow: 'general' (no faces) can unlock lower prices; 'mature' routes only to sellers without content restrictions.general · people · mature | people |
| callbackUrl | string (https URL) | Optional. When the run finishes (completed or failed), we POST the final run status to this URL — same shape as the poll response, plus model, params, and chargedCredits. Best-effort; polling remains the source of truth. | |
Unknown fields are rejected with a helpful 400. Media fields take https URLs. There is no provider parameter — every run is routed to the cheapest capable provider automatically.
Good to know
Billing
Runs bill your account credits at provider cost with 0% markup. The submit response includes the credit estimate; failed generations are refunded automatically. Add ?dryRun=1 to get the estimate without running anything.
Errors
Errors are always { "error": { "code", "message" } } with stable codes: unauthorized, model_not_available, invalid_parameter, insufficient_credits, rate_limited. A failed run is not an HTTP error — poll responses report status: "FAILED" with the reason.
What makes Seedance 2.5 Image to Video different
Up to 30 seconds in a single generation
Most AI video models cap out at 5-15 seconds. Seedance 2.5 generates up to 30 seconds in one pass — enough for a complete scene with a beginning, middle, and end. No stitching, no continuity drift between clips, no re-prompting to extend. Pick any whole-second duration from 4 to 30.
Synchronized audio, on by default
Every generation ships with native audio: spoken dialogue, ambient sound effects, and background music that match what's on screen. There is no separate audio pass and no lip-sync post-processing — the model generates picture and sound together. Turn it off with a single toggle if you want silent footage.
Built for editing pipelines
Alongside standard mp4, Seedance 2.5 can output mov encoded as yuv444p — higher color fidelity that holds up through multi-round editing and extension workflows where recompression loss normally accumulates. Feed the output back into an edit, upscale, or sequence node on your Scenetra board without visible degradation.
One model, three modes
Text to Video is one of three Seedance 2.5 modes on Scenetra. Start from a first-frame image with Image to Video, or drive generations from up to 30 reference images, 10 videos, and 10 audio tracks with Reference to Video — including audio-only referencing, where a single music track drives pacing and lip-sync.
More made with Seedance 2.5 Image to Video
Parameters
| Parameter | Type | Description | Default |
|---|---|---|---|
| Prompt* | text | Text prompt describing the desired video. Supports English and Chinese. | — |
| Duration | select | Video duration from 4 to 30 seconds. | 5 |
| Resolution | select | Output video resolution.480p · 720p | 720p |
| Aspect Ratio | select | Video aspect ratio, or 'adaptive' to select automatically from prompt context.16:9 · 4:3 · 1:1 · 3:4 · 9:16 · 21:9 · adaptive | adaptive |
| Output Format | select | mp4 by default; mov encodes yuv444p for higher color fidelity in multi-round editing pipelines.mp4 · mov | mp4 |
| Generate Audio | select | Synchronized audio — voice, sound effects, and background music — generated with the video.false · true | true |
Pricing
From $0.14 per second, audio included
Pay per generation — only pay for what you use.
480p
$0.14/s
per second
720p
$0.30/s
per second
What a video costs
| Duration | 480p | 720p |
|---|---|---|
| 5 seconds | $0.70 | $1.51 |
| 10 seconds | $1.40 | $3.02 |
| 30 seconds | $4.20 | $9.03 |
Use cases
Film & storytelling
- Complete 30-second scenes
- Dialogue clips with lip-sync
- Establishing shots
- Previz and animatics
Marketing & ads
- Product spots with music
- UGC-style ads
- Brand stings
- Concept testing
Social content
- Vertical 9:16 stories
- 30-second narratives
- Sound-on hooks
- Series with consistent style
E-commerce
- Product showcase clips
- Lifestyle b-roll
- Seasonal campaign variants
- Listing videos
Frequently asked questions
What is Seedance 2.5 Image to Video?+
It's the image-to-video mode of ByteDance's Seedance 2.5 model: upload an image and it becomes the first frame of a generated video up to 30 seconds long, with synchronized audio created alongside the picture. On Scenetra it runs as a node in the visual workflow editor — no API key needed.
How does first and last frame control work?+
The required image sets the video's first frame. Optionally add a second image as the last frame, and the model generates the motion that connects them — useful for transformations, camera moves that land on a specific composition, and seamless loops (use the same image for both ends).
What image formats and sizes are accepted?+
JPEG, PNG, WebP, BMP, TIFF, and GIF, with dimensions between 300 and 6000 pixels per side, aspect ratio between 0.4 and 2.5 (width/height), and file size up to 30MB.
Does the video keep my image's aspect ratio?+
Yes — Seedance 2.5 image-to-video always preserves the source image's aspect ratio (the mode is 'adaptive' by definition). Feed it a 9:16 portrait and you get a 9:16 video; there's no cropping or letterboxing to a fixed frame.
Does image to video generate audio too?+
Yes — audio is on by default in every Seedance 2.5 mode. The model generates ambient sound, effects, and music that match the animated scene. You can toggle it off for silent output.
When should I use Image to Video vs Text to Video?+
Use Image to Video when you need to control exactly what the first frame looks like — a product shot, a character design, generated art you want to keep. Use Text to Video when you're starting from an idea and want the model to invent the visuals. Both cost the same and support the same durations, so it's purely about how much visual control you need.
Related models
Seedance 2
ByteDance's Seedance 2.0 generates cinematic videos from text prompts with synchronized audio — voice, sound effects, and background music. Pick any duration from 4 to 15 seconds and resolutions up to 1080p, with per-second pricing that keeps 480p drafting cheap.
View model →Seedance 2 Mini
ByteDance's Seedance 2 Mini is the fast, lower-cost tier of Seedance 2 — generating cinematic videos from text prompts with synchronized audio. Mini delivers the same audio-visual sync at a fraction of the price, ideal for rapid iteration and high-volume social content.
View model →Kling 3.0 Pro
Kuaishou's flagship video generation model delivering stunning visual fidelity. Kling 3.0 Pro supports text-to-video, image-to-video with start/end frames, element references for character consistency, and native audio generation.
View model →Start creating with Seedance 2.5 Image to Video
Use Seedance 2.5 Image to Video alongside 50+ other AI models in Scenetra's visual workflow editor. No setup required.
Get Started