MiniMax H3 (Hailuo 3) is live on EvoLinkTry it with 10 free credits
Coming soon — not publicly available yet

Seedance 2.5 — AI Video Generator

ByteDance's next-gen video model: 30-second clips, up to 50 reference inputs, and region-level edits. Preview it here — and generate with Seedance 2.0 while you wait.

Runs on Seedance 2.0

Seedance 2.5 isn't publicly available yet — your prompt runs on Seedance 2.0 today.

Sample videos with full prompts

Seedance 2.5 preview footage — every clip ships with the exact prompt and settings behind it. Copy them, adapt them, make them yours.

One-shot 30s story: a football across 3,000 years

Produce a 30-second short popular science video about the three thousand years of evolution of football. The entire film uses the same ball as the main visual line. The ball rolls, travels, and deforms from ancient times, connecting different civilizations and eras. The overall rhythm is compact, the graphics are high-end, the historical science short film is combined with artistic transitions to highlight the feeling of a ball spanning three thousand years, and the voiceover is simple and powerful. At the beginning, an ancient ball slowly appeared from the black background with a texture of time on the surface, and then rolled into a Cuju scene from the Warring States Period in China. The picture turned into an ink style, referencing the style of <<<image_1_1>>>. Ancient people dressed in ancient costumes played Cuju in the courtyard with elegant movements and the ball bounced under their feet. Voiceover: The football story starts with Cuju. Then, the ball continues to roll forward, and the picture naturally transitions to an ancient Greek ball game scene. The picture is in the style of a classical oil painting, and the style refers to <<<image_2_2>>>. The background of the square and stone pillars is obvious, and people wearing ancient Greek robes play football. The picture is thick and historical. Voiceover: The Greeks loved ball games too. Then the ball rolled into medieval Europe, and the picture still maintained the style of oil painting. Villages, mud fields, and ordinary people chased the ball. The atmosphere was warm and rough, just like ancient folk football continuing the fire. Voiceover: European folk football keeps the game alive. Then the ball was kicked out, and the screen switched to a black and white documentary style. Referring to <<<image_3_3>>>, the scene arrives in 1863 England. Gentlemen, clubs, and grass pitches gradually appeared, symbolizing the official birth of modern football. This ball showed the standard appearance of modern football for the first time. Voiceover: 1863 — modern football takes shape. Then the scene quickly enters the modern era, with the ball spinning in the air, bringing out key development nodes in turn. The lights, stadium, spectators, trophies, and different scenes around the world are intertwined, showing that football has evolved from a local sport to a global one. At the end, the ball is in the center of a modern stadium, and crowds and cheers from all over the world merge in the background, forming the feeling of "a ball connecting the world". The picture is grand and epic. Voiceover: Now, football connects the whole world.

720p · 31s · native audio · 3 reference images · translated prompt

17 reference inputs, one continuous shot

Core instructions: A 26-second one-shot narrative short film, with stable tracking and interweaving, refer to <<<video_1_1>>>, and smooth surround movement, refer to <<<video_2_2>>>. Smooth travel feeling. The alternation of day and night and the flow of the four seasons are realized within the lens. The protagonist is a European woman <<<image_1_3>>>, placed in a sea of people full of fireworks, highlighting the ultimate sense of loneliness and the quality of cinematography. Segmented camera movement and scene description: 0-3 seconds (smooth back): The old wooden door <<<image_2_4>>> opens with a creak, and the camera follows the figure of a European woman wearing <<<image_3_5>>> walking out. She pauses slightly at the threshold. The streets ahead were mottled with light and shadow, and the sounds of hawking and crowds were approaching. She looked distant and slowly walked into the street. 3-6 seconds (back and side tracking shot): The camera keeps a smooth follow-up, she walks into the crowded morning market, the atmosphere refers to <<<video_3_6>>>. The two sides were crowded with colorful fruit stalls and spice shops, and a group of street jugglers were breathing fire dragons, refer to <<<image_4_7>>>. The firelight illuminated the crowd, but she did not squint and walked through at a steady pace. 6-9 seconds (Smooth Surround from the Side): The camera begins to smoothly wrap around to the side and front, capturing her profile. She walked past the noisy butcher's shop <<<image_5_8>>>, and a young mother passes by her holding a baby <<<image_6_9>>>. The baby stared at her curiously, but she just lowered her eyes slightly to avoid looking, without stopping at all. 9-12 seconds (forward and backward follow-up shooting): The camera continues to circle directly in front of the protagonist and performs backward and follow-up shooting. The crowd in front suddenly retreated to both sides naturally like Moses parting the sea. A huge elephant <<<image_7_10>>> covered in gorgeous red cloth appeared from the right side of the screen with a steady pace, occupying most of the screen. 12-15 seconds (gap penetration and rewinding): At the moment when the woman and the elephant are about to collide, the camera cleverly slides through the narrow gap between the elephant and the woman, recirculating back to her back. The elephants passed by hugely and silently, and the urchins chased them with joy. Like bells and laughter, she didn't even slow down. 15-18 seconds (ambient light and shadow gradient): As she walks, the light and shadow in the long shot change magically - the dazzling sunlight in midsummer softens instantly, a breeze rolls up the golden leaves <<<image_8_11>>> in the sky, and the season seamlessly transitions to late autumn in the same long shot. Fallen leaves brushed her shoulders. 18-21 seconds (360-degree immersive surround): The front suddenly falls into a grand street celebration <<<image_9_12>>>. Colorful ribbons and shredded paper burst into the air, and vendors leaned out to cheer. At this moment, the camera unfolds a continuous 360-degree panning movement, creating an extremely strong visual tear between the quiet and lonely protagonist and the frenzied surroundings. 21-24 seconds (circling back to the side and back): When the camera circled around and returned to her side and back, the falling ribbons had quietly turned into snow all over the sky - it was winter in an instant <<<image_10_13>>>. Pedestrians held up umbrellas or put on hoods, and the woman hunches slightly, turns up her coat collar, changes into <<<image_11_14>>>, and continued to walk alone in the snow. 24-26 seconds (slow push and beat): As she walked towards the end of the long street, the sky darkened at a speed visible to the naked eye, and the day sank seamlessly into the night. The dim street lights on both sides and the light bulbs of the stalls were turned on one after another <<<image_12_15>>>. The vendors were packing up their goods. The noise seemed to be slowly absorbed and distant by the heavy snow, and her steps gradually slowed down. Grand fireworks suddenly bloomed in the night sky <<<image_13_16>>>, the sound of fireworks blooming refers to <<<audio_1_17>>>. Colorful light spots flickered and jumped on the building walls and in her eyes. The world is still lively, but she looks up quietly, the camera slowly zooms out, and ends here gently.

720p · 27s · native audio · 17 reference inputs · translated prompt

Region edit: same duel, new castle, new costumes

Replaced the original version of the two-person martial arts video <<<video_1_1>>> with an empty-handed warm-up before a cold-weapon duel. The scene is replaced with a medieval stone castle platform, an ancient courtyard flat, a mountain fortress outer platform, or a simple stone brick duel field. The background is the ancient castle wall, wind, fog, distant mountain line, and the ground is flat and stone <<<image_1_2>>>. The clothes of the man in dark clothes in the video are replaced by <<<image_2_3>>>, and the clothes of the man in light clothes in the video are replaced by <<<image_3_4>>>. The action remains the same, without changing the original rhythm. AI special effects only enhance the environment and texture: wind blown clothes, light fog, a small amount of dust at contact points, metallic cold reflective texture, slight particles and epic color palette. The overall style is restrained, realistic, and a classic hardcore duel atmosphere. Background music synced to the beats.

1112×834 · 8s · native audio · 4 reference inputs · translated prompt

FPV one-shot: hello in 11 languages

One-shot FPV drone first-person video, 33 seconds of continuous long shot, no editing, no jump cuts, no transitions. The camera starts from inside the high-altitude clouds and forms a continuous descending flight line along the clouds, fog, light and shadow, valleys, waterfalls, lakes, flower fields, urban buildings and near-ground squares. 11 clear and independent language display blocks appear in sequence throughout the process. Each block only displays the text corresponding to a single language. There is no mixing, overlapping, or adding other languages. 0–3 seconds,<<<image_1_1>>> clouds naturally form “Hello” in Chinese; 3–6 seconds,<<<image_2_2>>> Mist and volumetric light form English “Hello”; 6–9 seconds,<<<image_3_3>>> High-altitude water vapor and sunlight projection form the Spanish (Mexico) "Hola"; 9–12 seconds,<<<image_4_4>>> Ribbon in the sky forming the Indonesian word “Halo”; 12–15 seconds,<<<image_5_5>>> kites formation forming “Hai” in Malay; 15–18 seconds,<<<image_6_6>>> Valley morning mist forms Thai “สวัสดี”; 18–21 seconds <<<image_7_7>>> Waterfall mist forming Arabic مرحبا 21-24 seconds,<<<image_8_8>>> The reflection on the lake and the ripples form the Portuguese word "Olá"; 24–27 seconds,<<<image_9_9>>> Flower fields and meadows are naturally arranged into Vietnamese “Xin chào”; 27–30 seconds,<<<image_10_10>>> City glass buildings reflect light and shadow to form the Japanese word "こんにちは"; 30–33 seconds,<<<image_11_11>>> Nearby fountain water mist, floor paving and light strips form the Korean word "안녕하세요". The overall atmosphere is an early morning sunrise, with golden backlight, soft volumetric light, real clouds and fog, natural motion blur, and movie-level realism. The camera speed starts slowly from 3–5 m/s, gradually accelerates to 14–16 m/s across the natural landscape, and then slows down to 2–3 m/s to hover stably in the near-Earth square. Lens parameters: wide-angle lens, 24fps, smooth FPV drone movement, pitch gradually transitions from -5° to -18°, and finally returns to 0°; slight yaw ±10°, roll controlled at 0–10°, ensuring a continuous, stable, and realistic sense of flight from shot to shot.

720p · 30s · native audio · 11 reference images · translated prompt

Cinematic brand film from a text prompt

[Overall style] A 30-second high-end brand-level visual blockbuster with a strong cinematic feel and high-end texture. The picture emphasizes dreamy light spots (Bokeh), silky motion blur transitions (Motion blur), volumetric lighting, and ultra-realistic material detail expression. [Storyboard] [0-5 seconds]: Dream prologue and macro close-up Extremely high quality macro close-ups. A slender hand stretched into the air, and its fingertips touched colorful spots as bright and twinkling as stars. With the flow of light and shadow, the scene seamlessly and smoothly transitions to an elegant woman wearing a pure white tulle skirt, who is playing a vintage piano intoxicatedly. Shallow depth of field, the background blurs into a beautiful blue-green tone. [5-15 seconds]: Then cut to a smooth follow shot: a woman wearing a French wide-brimmed straw hat and a flowing white skirt, running lightly in the dense flower path full of pink-orange roses and blue hydrangeas. The light casts dappled light and shadow through the leaves, perfectly showing the gentle breeze, the realistic physics of the skirt fabric fluttering, and the ultra-realistic texture of the petals. [15-24 seconds]: Quiet aesthetics and light and shadow refraction The rhythm of the camera slows down and enters the ultimate beautiful slow motion (Slow-motion). A girl is sitting at a black wrought iron table next to a European retro golden fountain and reading quietly. There are crystal clear soap bubbles floating in the air, and the surface of the bubbles perfectly reflects the surrounding flowers and warm sunshine. The moment when water droplets splash is clearly visible, demonstrating the model's top-level rendering capabilities for transparent materials, water refraction, and complex lighting.

720×960 · 24s · native audio · text-to-video · translated prompt

Anime racing short, storyboarded second by second

30 seconds cinematic youth racing short film, 2d animation style. The protagonist is a young driver who drives a motorcycle to participate in high-profile competitions. The overall style is passionate, youthful, emotionally intense, and cinematic, with a complete beginning, transition, and clear emotional arc. Only two types of camera movements are used in the entire film: high-speed tracking and slow-motion surround. There are very few lines, and they appear naturally like memory fragments. The tone is sincere, gentle, and restrained, without shouting slogans or overly sensationalizing. Don’t feel disaster, don’t express negatively, and don’t exaggerate science fiction. It focuses on love, support, counterattack and growth in the race of youth. 0 seconds to 5 seconds The track begins at dusk with high-speed and intense racing. The camera follows the young man's motorcycle closely to the ground at high speed. The tires skim the edge of the track. The motorcycle roars, the wind blows fiercely, and the atmosphere is tense and fiery. The young man was concentrating, and the setting sun drew sharp highlights on the metal shell of the car. 5 seconds to 9 seconds After entering a key corner, the boy was suddenly overtaken by his opponent. The high-speed tracking continued, and the picture showed the oppressive feeling of the ranking declining and the rhythm being disrupted. In the close-up view of the helmet, there is temporary loss of concentration, tightness of breathing, and slight shaking. The boy whispered: "Can I still catch up..." 9 seconds to 14 seconds The boy fell behind, his breathing became heavier, and his mood hit a low point. The race did not stop and the locomotive was still moving forward at high speed. The scene began to flash back warm memory fragments during high-speed riding: when he was learning to drive as a child, someone supported him from behind; his father arranged his helmet for him, his movements were meticulous and quiet; before the finish line, a gentle smile looked at him; and the back figures walking side by side on the slope at dusk. These memories are presented with golden backlighting, soft slow motion, and fragmented feelings. 14 seconds to 18 seconds The music gradually turns from depressive to uplifting. A restrained and gentle voice came from the memory: "Don't be afraid — I'm always here." "Stay steady. And look forward." The young man's eyes refocused, his breathing slowly stabilized, and his mood changed from wavering to firm. 18 seconds to 23 seconds The young man regained his confidence, accelerated with all his strength, and counterattacked accurately. The high-speed tracking shots show the power and control of the motorcycle when cornering, exiting corners, and approaching the car in front. The boy said quietly but firmly: "I won't stop here." 23 seconds to 27 seconds An ascending track appeared ahead, and the boy sprinted at full speed against the setting sun. The picture only retains the sound of breathing, engine sounds and continuously rising music, without adding unnecessary lines. The locomotive soared into the air with the help of inertia and entered shocking slow motion. A gentle voice with a smile came from deep in my memory: "Go on." 27 seconds to 30 seconds The camera circles the locomotive in mid-air for a slow-motion panoramic close-up. Push the emotions of passion, tenderness, freedom and upward leap to the climax. Flowers bloom behind you, followed by seedance, refer to <<<image_1_1>>>

720p · 30s · native audio · 1 reference image · translated prompt

What is Seedance 2.5?

Seedance 2.5 is the next generation of ByteDance's Seedance video model, unveiled at the FORCE conference in Beijing in June 2026 as the successor to Seedance 2.0 — the model that currently leads Artificial Analysis' image-to-video arena.

As of this writing, Seedance 2.5 has not been publicly released. ByteDance's own Dreamina page still lists it as "coming soon", and access is limited to closed enterprise testing. Any third-party site claiming you can generate with Seedance 2.5 today is running a different model under the hood.

Once it ships, ByteDance has said Seedance 2.5 will power its consumer apps — Jimeng and Doubao in China, Dreamina and CapCut internationally.

What's new vs Seedance 2.0

  • 30-second single-shot clips, up from 15 seconds — with a beta long-video mode that extends to 180 seconds (official announcement).
  • Up to 50 multimodal reference inputs — text, images, video, audio and style — up from Seedance 2.0's cap of 9 images, 3 videos and 3 audio clips (official announcement).
  • Region-level editing: change a specific area of a generated video without re-rendering the whole clip (official announcement).

Try it in three steps

1

Open the playground

Seedance 2.5 is still in closed testing, so the playground runs Seedance 2.0 — same family, available right now. New accounts get 10 free credits.

2

Write or paste a prompt

Start from a sample prompt on this page or describe your own scene: subject, action, camera, style.

3

Generate and iterate

Review the result, tweak your prompt, and re-run. We'll switch this page to Seedance 2.5 the day it opens.

Capabilities and limits

What ByteDance says it can do

  • 30-second clips in a single generation, extendable to 180 seconds in a beta long-video mode (official claim).
  • Up to 50 reference inputs across text, image, video and audio for character and style consistency (official claim).
  • Region-level editing without regenerating the full clip (official claim).
  • Audio as a supported reference input, building on Seedance 2.0's native synced audio generation (official claim).

What to know before you plan a project

  • Not publicly released: no consumer app access yet. Third-party sites advertising instant Seedance 2.5 access are running other models.
  • Early testers with preview access report 720p output so far, while official materials promise 4K — treat resolution claims with caution until wide release.
  • Several key specs, including frame rate, remain officially unannounced.
  • Seedance 2.5's face-reference policy is unannounced — real-face support on this page currently runs on Seedance 2.0.

Shipping video generation inside your own product? The developer page tracks Seedance 2.5 integration status and everything technical.

Seedance 2.5 API

Frequently asked questions

Everything you need to know about the product and billing.

Not yet: Seedance 2.5 is not publicly available. ByteDance's own Dreamina page lists it as "coming soon", and access is limited to closed enterprise testing. Third-party sites advertising instant Seedance 2.5 access are running other models. On this page you can try Seedance 2.0, the current publicly available generation, in the playground.
Seedance 2.5 is ByteDance's next-generation AI video model, unveiled at the FORCE conference in June 2026 as the successor to Seedance 2.0. Its three headline upgrades: 30-second single-shot clips, up to 50 multimodal reference inputs, and region-level editing of generated video.
Three official upgrades: clip length grows from 15 to 30 seconds (with a 180-second beta long-video mode), reference inputs grow from 9 images + 3 videos + 3 audio clips to 50 across all types, and region-level editing lets you change part of a clip without regenerating it. See our full comparison for details.
ByteDance hasn't announced how Seedance 2.5 will be offered or what it will cost. Treat any published numbers as guesses. New accounts on EvoLink get 10 free credits to try Seedance 2.0 in the playground.
Yes — every clip in the gallery above was generated with Seedance 2.5, and each ships with the exact prompt used to create it, shown in faithful English translation, so you can judge — and reproduce — what the model actually does.
Yes — the playground on this page supports real-face references with Seedance 2.0. Upload only faces you have permission to use. ByteDance hasn't published Seedance 2.5's face policy yet — we'll update this answer when it does.
Official preview cases include a full 2D-anime racing short — it's in the gallery above with its prompt. Independent hands-on testing isn't public yet; anime is one of the first things we plan to test when access opens.
Audio is listed as a supported reference input for Seedance 2.5. Its predecessor generates native synced audio — dialogue, effects and music — and ByteDance hasn't separately detailed 2.5's audio output yet.

Explore the Seedance family

Last updated: