Seedance 2.5 is live on Skiprr
ByteDance's new flagship video model shipped today, and it is already the default engine for every Skiprr video production. One-take 30 second stories, reference sets of up to 30 images, and native synced audio, ready in your workspace.
ByteDance released Seedance 2.5, the new generation of its video creation model, and Skiprr adopted it the day it landed. Every new video production in Skiprr now runs on Seedance 2.5 by default. No keys, no setup, no waitlist. You describe the video, attach your references, and the strongest video model available today does the rest, with synchronized dialogue, sound effects, and music generated in the same pass as the picture.
Seedance 2.5 centers on two things that matter for real production work: foundational generation that holds together over long takes, and reference based generation that follows your materials instead of improvising around them. Three upgrades carry the release.
Thirty seconds. One take. A real story.
Single pass generation doubles from 15 to 30 seconds, and the extra room is used for structure, not padding. Within one take the model organizes multiple logically connected shots so a story unfolds through setup, development, turn, and resolution. Shot transitions, subject stability, and audio sync all hold across the cut points, so long clips read like coverage, not like a loop.
T2V promptOne-take handheld gimbal tracking shot. The camera slowly pushes in through a gap in a heavy red curtain and enters a warm-toned backstage dressing room. A young female singer, with her back to the camera, is adjusting her earpiece as a staff member reminds her it's time to go on. She turns toward the camera and starts singing citypop. The camera pulls back and tracks her as she passes through the curtain into a dim backstage corridor, interacting naturally with her dancers along the way; one staff member hands her a microphone. She and the dancers then step onto the stage, and the camera arcs around to the back, gradually revealing the red-and-black stage design, LED screens, spotlights, haze, and reflective floor. The camera finally pulls out to a wide shot of the arena, showing the packed audience, light boards, glow sticks, and cheering crowd, capturing the youthful, free-spirited climax of the concert.
Multi round extension continues an existing output while keeping characters, environment, and pacing consistent, so several 30 second passes chain into minutes of continuous story without splicing or fixing transitions by hand.
R2V promptExtend the video. Continue from the visuals and subjects in @Video 1 and generate another 30-second clip, keeping the character subjects, scene, visual style, and sound effects consistent. The little boy runs along the train carriage holding a soccer ball. When the subway stops, the side door opens and he immediately dashes out, with the male lead chasing after him. The two run across the platform and out onto the street, startling passersby and vehicles along the way. The male lead finally catches up and grabs him. The boy looks up, aggrieved. The male lead's anger slowly fades; he pats the boy's head and shows a helpless smile.
Referencing at production scale
A single generation now accepts up to 30 reference images, 10 video clips, and 10 audio clips. The model reads composition, scene, style, characters, and props across all of them and applies each one where your prompt says to. Multi character shots keep every face, wardrobe, and voice stable. Clay render references let you block a scene in untextured 3D, and the model honors your camera moves, motion paths, and staging while lighting the final render to physical rules.
R2V promptA 30-second concert sequence in 16:9 landscape, with cinematic realism, authentic concert hall lighting and shadows, warm golden stage lighting, and the atmosphere of a formal classical concert. Use @Image 1 for the venue. Reference @Image 2 for the pianist. Reference @Image 3 for the cello. Reference @Image 4 for the violin. The lead vocalist must strictly follow @Image 5. Reference @Images 6 to 10 for the rest of the orchestra. Reference @Images 11 to 14 for the choir. Reference @Images 15 to 18 for the audience seating. The lead vocalist walks from center stage toward the front edge. The pianist is positioned by the piano. The orchestra is arranged on both sides and toward the rear. The choir stands at the back of the stage. Open with a high-angle wide shot of the full concert hall. The pianist strikes the keys, and the lead vocalist steps into the spotlight and begins singing. The camera naturally moves across the violin, cello, and orchestra as they perform together, with the violin feeling bright and the cello warm. In the latter part, the choir joins in. The lead vocalist briefly makes eye contact with front-row audience members, who respond with a smile and a slight nod. In the closing shot, the camera pulls back. The singing ends, and the audience joins in the applause.
R2V promptRefer to @Clay Render 1 for camera movement, pacing, shot-size transitions, subject trajectory, and blocking. Refer to @Image 2 for character design, scene, materials, lighting, color, and fairy-tale atmosphere, and render the white model as a dreamy, warm, 3D animated short with a childlike fantasy feel. The story unfolds as follows: flight through a fantasy sky, mythical beasts flying alongside through a sea of clouds, a dive into the ocean, weaving through the deep with manta rays, passing through a mirrored rift in spacetime, picking stars from the cosmos, transforming back into the bedroom, father tucking in the blanket, the picture book closes and holds on the final frame.
Editing you can direct to the second
Seedance 2.5 takes timestamp level direction. In one prompt you can control narrative, camera, and rhythm for each time range, and after generation you can modify characters, actions, or plot inside a specific window while everything around the edit stays continuous. Green screen editing swaps the world around a subject and re renders how clothes, hair, gait, and lighting respond to the new environment. Camera perspective editing re shoots the same scene with a new camera plan while the action stays untouched.
R2V promptUsing @Video 1, render the green-screen background, obstacles, wardrobe, and supporting characters. 0-4s: outdoor training, replace the obstacles with rocks, bricks, tires, and wooden crates. 4-10s: locker room, friends offering encouragement. 10-15s: international match, replace the training poles with original defenders and a goalkeeper, and the protagonist scores. Overall photorealistic, cinematic quality.
R2V promptEdit @Video 1. Keep the characters, actions, and visual style unchanged. Adjust only the camera movement. A 15-second segmented camera plan: 0-4s, a micro-FPV move skims tightly past the pan, then follows the popping toast and whip-pans to the coffee; 4-7s, push in and track laterally along the rim of the pan, following the fried egg as it flips up and lands back in place; 7-11s, rapidly rise to a top-down view, then descend at a steady pace, sweeping across the plate and keys; 11-15s, use a handheld close-up to follow the hands with a fast lateral whip, then push in on the breakfast and pull back to a medium two-shot. Keep the entire sequence smooth, continuous, and stable.
Beyond the studio
The same capabilities are reaching classrooms and factory floors. Teachers turn historical context and abstract principles into vivid demonstrations. Industrial teams generate synthetic training footage, process walkthroughs, and equipment demos from clay renders and a single style reference.
Already the default in Skiprr
Open a video production and Seedance 2.5 is preselected. Work in text mode for pure prompts, image mode to animate a start frame, or reference mode to attach the images, clips, and audio the video should follow, then cite each one right in the prompt as [Image1], [Video1], or [Audio1]. Dialogue goes in double quotes and comes back lip synced. Duration runs up to 30 seconds per clip, or let the model pick the length that fits the story.
Everything is billed in the same credits as the rest of your workspace, with live per second pricing and an estimate before you launch. Your brand brain still runs the show: references, brand voice, and guardrails carry into every generation, so the newest model in the world still sounds like you.