Seedance 2.5 video prompts, each shown next to the clip it produced and credited to its author. Where the creator published their prompt we keep it verbatim; where they did not, we reverse-engineer one from the finished clip and label it as such. Hover to play, one click to copy.
Seedance 2.5 is ByteDance's video model, unveiled on 23 June 2026 at the Volcano Engine FORCE conference. On the shipped API it generates 4 to 30 seconds in a single pass — or duration -1 to let the model choose the length — and accepts up to 50 reference assets in one job: up to 30 images at 4K each, up to 10 video clips and up to 10 audio clips totalling 30 seconds each. It adds video editing and video extension as task types, and generates audio in the same latent space as the picture. Resolution is 480p or 720p; the native 4K at 10-bit shown at the FORCE launch is not exposed on the API. ByteDance also published an official prompting guide for it. This page collects Seedance 2.5 video prompts with the clip each one produced, free to copy. Since 8 August 2026 you can run Seedance 2.5 on apimodels.app as the model seedance-2.5, from $0.134/s; the creators in this library ran these prompts in Dreamina (Jimeng) and Higgsfield.
Source: ByteDance Seed (the model vendor’s official site)
Hover to preview · click for sound
Red-hooded armoured soldier walking out of desert ruins
Reconstructed
@image1 supplies <SOLDIER>'s armour design, colour breakdown and silhouette: chipped bone-white and gunmetal plating, a heavy respirator faceplate with a single narrow red visor slit, and a sun-bleached crimson hood and cape frayed at the hem. Do not take the background of the reference; do not take any pose from it.
<SOLDIER> walks alone through a bombed-out mud-brick settlement at the flat hour before noon, crossing open sand, passing between collapsed walls, and continuing down the main street without stopping.
The image is anamorphic live-action science fiction: heavy suspended dust, ochre and bleached-bone grade, hard overhead sun with deep contact shadows, sand grain caught in the armour joints, real cloth weight in the cape, and a shallow depth of field that keeps the ruins soft.
The camera works in five moves. It opens at ground level behind <SOLDIER>'s boots as sand collapses under each step and the cape drags through frame. It cuts to a slow push-in on the faceplate, the red visor the only saturated colour in the frame. It tracks laterally past a broken wall so the wall wipes the lens. It follows over the shoulder down the ruined street with the horizon shimmering. It ends on a locked-off wide, <SOLDIER> a black silhouette walking into a blown-out sun at the end of the street.
Keep the armour damage pattern, hood shape, cape length and the position of the visor slit identical in every shot. Keep one figure in frame at all times.
Sound includes wind across open sand, grit under armoured boots, servo movement in the joints, the low filtered rasp of breathing through the respirator, distant loose metal, and a single sustained low drone with no melody.
Fourteen images and an audio track into one music video
Create a fast, trippy 30-second dubstep music video using all the uploaded photographs
(PLEASE FIND ATTACHED ALL 14 IMAGES + AUDIO IN THREAD).
Use the uploaded 30-second audio as the only soundtrack. Make it a powerful dubstep edit: every movement, punch-in, transition and hard cut must hit the beat. Fully animate everything inside the photographs. Make the dancers perform, fabric and flames spin, painted eyes react, masks lunge, multiple arms ripple, the drummer strike his drum, sadhus dissolve into ash and black birds, motorcycles race aggressively, smoke creatures come alive, the underwater band actively play, boats move through the surreal landscape, and the flaming weapon create violent circles of fire. No photograph should remain static.Freely mix all the visuals in whatever order works best with the music. Let movement from one photograph naturally transform into the next: spinning fire becomes fabric, fabric becomes a painted eye, expanding arms become the drummer’s arms, ash birds become racing motorcycles, motorcycle dust becomes dancing smoke, smoke becomes underwater bubbles, bubbles become water and boats, and the boats transform into circles of fire.Keep constant movement throughout. Use rapid punch-ins, speed ramps, bass-impact flashes, hard rhythmic cuts, brief freeze-frames and fluid psychedelic transitions. Add a separate graphic layer flashing ॐ, the trishul, damru, third eye, crescent moon and traditional Hindu sacred symbols precisely on the beat. Let them pulse, rotate, distort and dissolve without covering the performers.Make it stylish, colourful, strange and extremely high-energy—an edgy, award-winning Incredible India advertisement edited like a dubstep music video. No dialogue, no typography, no slideshow, no long static shots and no complicated story. Finish with the strongest collision of fire, ash, birds and sacred symbols on the final bass hit.
#capcutseedance25 #capcutcpp
One image overrides every earlier performer reference
@Image1 is the absolute reference for THE RAPPER and completely replaces every previous performer reference. Preserve his exact identity: middle-aged man with a high receding hairline, short salt-and-pepper hair, thick dark eyebrows, dark eyes and a full beard with strongly defined white-gray sections. Preserve his stocky build, black-white-dark-green horizontally striped T-shirt with black chest pocket, sand-colored knee-length chino shorts and chunky off-white sneakers. No changes to his face, body, hair, beard, clothes or proportions.
A 30-second single continuous live stadium rap performance captured horizontally on an iPhone from the front audience section. Authentic handheld fan footage: physical hand tremor, operator breathing, imperfect reframing, rolling shutter, digital-zoom softness, momentary autofocus hunting and compressed phone-microphone sound. No cuts.
The first frame already shows THE RAPPER full-body on the right third at the end of a stage runway. A huge sold-out stadium surrounds him. Exactly four adult backup dancers wait several meters behind him. Emerald, white and black LED graphics echo the stripes of his shirt. The stage floor remains solid, flat and continuous.
A heavy original grime beat begins: deep sub-bass, dry kick, snapping snare and minimal low synth. The iPhone rapidly pinches from 1× to a shaky 5× digital zoom, briefly overshoots, then locks onto THE RAPPER in a full-body composition. Focus stays wide enough to preserve his feet and choreography.
He begins rapping with a low-mid, forceful cadence and exact lip synchronization:
THE RAPPER:
“Walk in steady, put the weight on the beat,
Every bar lands, every move stays clean.
Hands up high when the bass comes down,
I don’t chase the wave—I shake the whole ground!”
Only these words are spoken. Each line is delivered in one controlled breath.
On the first bar he performs two violent shoulder hits, a chest pop and a sharp forearm lock. His shirt and beard react naturally to momentum.
On the second bar he executes fast heel-toe pivots, crosses one foot behind the other and glides sideways while keeping his heavy body convincingly grounded. The camera operator struggles to keep his sneakers in frame, corrects downward and catches the complete footwork.
On the third bar the four dancers join in perfect synchronization. THE RAPPER leads a hard sequence: right stomp, left stomp, elbows strike outward, torso snaps backward, hands shoot overhead. Every movement lands precisely on a kick or snare.
The instrumental cuts for one beat. He holds a deep wide stance, eyes fixed on the upper tiers. His chest rises with one visible breath.
He shouts the final line while performing a rapid three-step, a controlled 180° pivot and one enormous downward arm strike. On “GROUND,” he stomps once. The bass returns with a massive impact; the LED floor sends a broad solid emerald light wave across the stage.
The entire stadium copies his movement. The phone shakes from thousands of spectators stomping together. He breaks into a confident grin but continues bouncing in time, pointing from one side of the stadium to the other.
The operator zooms rapidly back through 3× and 1× to 0.5× ultra-wide while turning 160° away from the stage in one continuous handheld sweep. Exposure briefly pumps, then recovers.
Finish on the full stadium bowl: tens of thousands of people across every tier performing the same shoulder-hit and stomp combination, emerald wrist lights moving in broad geometric waves, stage remaining on the far-left edge. Audio continues with the crowd chanting “SHAKE THE GROUND,” live bass vibration and realistic phone compression. Rich emerald and white light, warm skin and deep blacks. Clear air without haze, smoke, confetti, mist, sparks or airborne particles.
image_1 is the first frame. The woman from
image_1 — straight black hair, white short-sleeve collared shirt, navy pleated skirt, brown belt, braided bracelet — is holding the camera operator's hand and pulling him forward through the world. Camera is first-person POV, loose handheld, her hand visible in the lower right of frame gripping the operator's wrist, pulling. She walks ahead, glancing back over her shoulder occasionally. She walks, then runs, then sprints as the scale of the world gets bigger and more overwhelming. Every few seconds a vertical floor-to-ceiling insert appears ahead floating in open air — no border, no frame, no door, no arch, just a hard rectangular cut in space where one world ends and another begins — and she pulls him straight through it into the next world. Six locations total, one continuous uncut shot. Location 1 — 1970s apartment party interior from image_1: warm amber light, sequin curtain, strangers in background, she turns and pulls him out of the room forward. Location 2 — Petra, Jordan, midday: the narrow Siq canyon walls rising hundreds of meters on both sides, sunlight cutting down in a single strip, she runs pulling him through the corridor of ancient rock. Location 3 — Shibuya crossing, Tokyo, rush hour: hundreds of people crossing in every direction, she weaves pulling him through the crowd, neon signs and screens overhead in daylight. Location 4 — Sahara desert at golden hour: vast open dunes, no horizon line visible, wind moving the sand, she runs pulling him up a dune face, her skirt and hair flying. Location 5 — Venice, Italy, acqua alta flooding: St Mark's Square underwater, ankle-deep reflective water, pigeons displaced onto benches, she splashes forward pulling him across the flooded piazza. Location 6 — top of a mountain at dawn, clouds below: nothing ahead but open sky and light. She slows to a walk. Then stops. Still holding his hand. She turns and looks directly into the camera — directly at him — for the first time. Silence. Sound: party ambience fading, then canyon echo of running footsteps, then Tokyo crowd noise and distant traffic, then open desert wind, then Venice water splash and bells, then pure silence and wind at altitude. Her breathing gets heavier location by location. Style: photorealistic cinematic, 2.39:1 Cinemascope, warm film grain consistent across all six spaces, first-person handheld POV, one unbroken take from start to finish.
Two character sheets, identity and wardrobe locked
Preserve the exact identity and wardrobe of the male target @target_character_sheet and the female assassin @assassin_character_sheet. Use @landcruiser_sheet as the identity reference for the black 1990s Toyota Land Cruiser. Use @kawasaki_ninja_sheet as the identity reference for the green Kawasaki Ninja motorcycle. Use @tokyo_street_mood_1990s and @chase_moodboard_1990s to ground the environment and chase tone. Cinematic anamorphic high-stakes movie chase scene. Begin in the middle of the chase with both vehicles already travelling at full speed through dense 1990s Tokyo traffic. The male target drives the black Toyota Land Cruiser. The female assassin follows closely behind on the green Kawasaki Ninja motorcycle. This is the climax of the chase. The Land Cruiser loses control in traffic, crashes violently, rolls, and explodes in a huge fireball. Because the assassin is following so closely behind, she is suddenly thrown into danger as well. She has only a split second to react. She swerves hard to avoid colliding with the crashing Land Cruiser and the flying debris, narrowly misses the wreck, briefly loses stability, then fights the motorcycle back under control and continues forward at speed. Create maximum perceived speed from the first frame. Use low road-level angles, front-quarter vehicle shots, wheel-level inserts, over-the-shoulder motorcycle pursuit views, compressed telephoto traffic shots, and one brief interior insert of the target reacting inside the Land Cruiser. The camera is repeatedly overtaken by the vehicles, falls behind them, then is rushed past as the crash unfolds. Traffic, barriers, road markings, lights, and water spray pass close to the lens, creating intense foreground parallax and rapid scale changes in frame. Build the scene as one escalating chain of high-speed pursuit, violent crash energy, rolling vehicle momentum, sparks, debris, smoke, and explosion. Keep the Land Cruiser heavy and physical. Keep the assassin fast, precise, and under real pressure as she reacts at the last possible moment, swerves through the danger, stabilises the motorcycle, and emerges clear of the wreck. Compose the action with strong cinematic geometry and visual hierarchy. Use opposing diagonals, asymmetrical subject placement, layered traffic, deep road perspective, compressed depth, and foreground wipes. Let the Land Cruiser dominate the frame during the crash, then shift the visual emphasis to the assassin and the motorcycle as she recovers control and becomes the clear focal point. End with a front-facing tracking shot of the assassin riding toward the camera at speed just after regaining control, still carrying the energy of the evasive move, with the huge explosion of the Land Cruiser behind her. Her posture should show intensity and effort, as if she has only just recovered from the near collision. Firelight, rain haze, neon spill, sodium streetlights, and glossy wet-road reflections. Cinematic anamorphic high-stakes movie chase climax. No music.
Use Image 1 for the large muscular cartoon bulldog boxer, Image 2 for the live-action Caucasian American boxer, and Image 3 for the vintage arena, ring, crowd, and lighting.
Create an intense 24-second boxing sequence in 10 shots.
The bulldog remains a hand-drawn 2D character inspired by 1969 American TV cartoons, keeping the same face, proportions, red gloves, and striped shorts. Never make him realistic. The human boxer remains fully photorealistic.
Style: heated live sports documentary broadcast. Handheld ringside cameras, reactive pans, imperfect reframing, small zoom corrections, natural shake, motion blur, focus breathing, faded 35mm colors, soft lens bloom, subtle gate weave, and naturally changing film grain. No frozen grain or fixed texture.
Strong white spotlights illuminate the ring. The crowd is darker but visible and constantly reacting. Crowd cheering continues from the first frame to the last: shouting, applause, whistles, and stomping. Reactions grow louder after major punches. An excited male English commentator speaks continuously. No music.
Each numbered shot is one continuous camera take. The camera may pan, track, dolly, or reframe, but must not jump to a new position inside the same shot.
Comic effects must resemble authentic 1969 American comics: imperfect hand-drawn lines, thick black outlines, red, yellow, white, and black colors, halftone dots, faded ink, and slight print misregistration.
Use effects sparingly:
SPEED LINES on fast punches.
One-frame SMEAR FRAME on the strongest punches.
Small IMPACT BURSTS on selected hits.
IMPACT FLASH and SHOCKWAVE RING only on the decisive counter and final uppercut.
WOBBLE LINES and yellow rotating stars only when the human is badly dazed.
Use only these comic sound effects on major clean hits:
THWACK! POW! SMACK! WHAM! CRACK!
SHOT 1 — 0.0–2.0
High-angle view of the packed arena. The ring is under strong spotlights. Both fighters wait in opposite corners. The bell rings, the crowd erupts, and both walk toward the center as the camera pushes closer.
SHOT 2 — 2.0–4.0
Slightly elevated wide shot from outside the ring with ropes in the foreground. The referee steps back. Both fighters raise their guards, circle, and test distance. The handheld camera tracks right.
SHOT 3 — 4.0–6.2
Low canvas-level angle. The bulldog throws a jab with brief SPEED LINES. The human narrowly avoids it and counters with a jab that the bulldog blocks. Quick natural pan. No comic text.
SHOT 4 — 6.2–8.8
Slightly high-angle shot through the ropes. The human attacks with a combination. A left jab lands with a small IMPACT BURST and THWACK! A right straight follows with POW!, a brief IMPACT FLASH, and a small burst. A left hook lands with only a subtle impact mark. The camera tracks as the bulldog retreats.
SHOT 5 — 8.8–11.2
Close handheld shot from the opposite side. The human keeps pressing with short punches. The bulldog blocks several and absorbs one body shot. The camera urgently reframes. The bulldog answers with a short face hook using a small red IMPACT BURST and SMACK!
SHOT 6 — 11.2–14.0
Over-the-shoulder from behind the human. The human throws a powerful right straight. The bulldog slips underneath. The camera pans with the missed glove. Without cutting, the bulldog plants his feet, rotates, and launches a right counterpunch with SPEED LINES and a one-frame SMEAR FRAME.
SHOT 7 — 14.0–16.8
Tight ringside impact shot. The bulldog’s right straight lands on the human’s cheek in very brief slow motion. His cheek compresses and sweat flies. At contact, show a large IMPACT BURST, white IMPACT FLASH, radial lines, short SHOCKWAVE RING, and WHAM! The effects vanish immediately. Return to normal speed as the human staggers backward. The crowd explodes.
SHOT 8 — 16.8–19.2
Slightly elevated ringside shot. The human is badly dazed but standing. WOBBLE LINES and three yellow stars briefly appear. The bulldog lands one body punch without text, then a short face hook with SPEED LINES, a small burst, and SMACK!
SHOT 9 — 19.2–21.5
Low diagonal ringside angle. The bulldog throws a compact uppercut with a one-frame SMEAR FRAME. It lands under the chin with IMPACT FLASH, red-yellow IMPACT BURST, short SHOCKWAVE RING, and CRACK! The human falls into the ropes, drops to one knee, and collapses safely. The referee stops the fight. The bulldog steps back. Yellow stars fade.
SHOT 10 — 21.5–24.0
Final slightly high-angle wide shot. The referee brings the bulldog to the center and raises one arm. Red-and-white paper streamers fly into the ring. Vintage photographers fire flashbulbs and strong white strobes. The crowd remains standing, shouting, clapping, whistling, and waving hats. The handheld camera slowly pushes closer until the final frame.41:T8ad,Style:
Authentic late-1970s Mediterr
The reference supplies identity only, nothing else
Prompt
IDENTITY: @Image1 supplies identity only — face, long black hair, low-brim cap, glossy black bodysuit, segmented cybernetic right arm, circuit tattoos, heeled boots. Frame is one continuous open-sky environment edge to edge. Cuts every 2-3s, no shot over 3s, one camera movement per shot. 0-2s | Worm-eye far below, whip tilt up as she pushes off the spire ledge. [Cut] 2-4s | Macro on her heel leaving the lip, static. [Cut] 4-6s | Top-down on her back, falling with her, hair rips straight up. [Cut] 6-8s | Boot macro, black plates streak in and latch over her shins, static. [Cut] 8-10s | Side track, thigh and hip plates snap shut. [Cut] 10-12s | Cybernetic arm macro, forearm plating extends over it, static. [Cut] 12-15s | Wide, chest harness clamps on, two stabilizer fins fold from her back, push-in. [Cut] 15-17s | Face close-up, a thin dark visor slides from the cap brim over her eyes, locked. [Cut] 17-20s | Extreme wide, geared figure punches the cloud layer, twin vapor spirals. [Cut] 20-23s | Head-down delta dive, fins biting air, hard push-in on her back. [Cut] 23-26s | Ground-level low angle, she swells in frame fast. [Cut] 26-28s | She lands in a deep crouch, cybernetic arm braced on stone, cracks radiate 4m in a spider-web pattern. [Cut] 28-30s | 0.4x slow motion, concentric shockwave ripple distorts the air outward, dust erupts to waist height, she rises into a recovery stance. GEAR: plates arrive as fast streaks from off-frame, each locking with a latch snap and a thin amber seam that fades in 0.4s and lights her skin. Assembly runs 6-17s, one region per shot. PHYSICS: terminal velocity, hair and fabric dragged upward, motion blur on background only, subject sharp, faster after 17s. AUDIO: wind roar climbing, latch snap on each lock, sub-bass on every cut, half beat of silence at 25.5s, air-pressure boom at 26s, rumble tail. STYLE: painterly anime illustration matching @Image1, cel-and-gradient shading, cool desaturated sky.45
Lock the original artwork completely. Preserve the exact illustration style, brushwork, textures, materials, color palette, lighting, creature proportions, facial expressions, environment, and character designs. Never change the art style or introduce new designs. Keep every dragon and child identical throughout. No text or watermark. One continuous forward-moving story. Begin with motion from frame one. Use cinematic hard cuts, speed ramps, whip pans, dynamic tracking, FPV fly-throughs, overhead drone shots, bullet time where appropriate, volumetric clouds, realistic fire physics, motion blur, depth of field, subtle film grain, debris, feathers, sparks, smoke, expressive facial animation, and seamless transitions.
0-2s
Cold open. The children happily eat while the giant dragon lazily sniffs the food. The baby dragons taste tiny bites, instantly make disgusted faces, stare at each other, then suddenly leap onto the table together. Fast push-in, comedic bass hit.
2-4s
The babies explode into chaos. Plates, bread and vegetables fly everywhere. One dragon flips a table with its tail, another steals a roast, another scatters books while the giant dragon watches with an embarrassed smile. Rapid whip pans and handheld camera.
4-6s
The giant dragon opens both wings. The babies launch from its back and burst through the castle ceiling into the sky. FPV camera follows them weaving between towers before soaring into bright clouds. Epic orchestral rise.
6-9s
High-speed aerial sequence. The baby dragons dive through clouds, barrel roll, narrowly miss cliffs, then suddenly spot a peaceful sheep meadow below. Bullet-time pause on their excited faces before they dive.
9-12s
They swoop low across the meadow. Tiny controlled bursts of fire instantly roast food over campfire-style racks beside the flock without harming people. The dragons eagerly grab the freshly cooked feast, stuffing themselves with exaggerated joyful expressions. Smoke, sparks, golden firelight and fast tracking shots.
12-15s
Now completely full, the dragons happily fly back through sunset clouds carrying roasted food. They glide into the castle and land beside the giant dragon. The children laugh instead of getting angry. The babies proudly share their food with everyone. The giant dragon smiles warmly as the camera cranes upward, ending on a cozy wide cinematic family shot.
Gothic-lolita twins, three references with distinct jobs
参照素材:
Image1: 銀髪ゴスロリ双子の顔立ち、目つき、髪型、衣装、黒い光輪、全体配色の最優先参照。元画像の構図は使用しない。
Image2: キャラクター設定画。正面・背面の髪型、体型、黒いゴスロリ衣装、靴、表情設計の参照。
Image3: 実写ゴシック寝室の舞台参照。部屋の構造、天蓋ベッド、レース、ベルベット、暗色木材、窓、ランプ、家具、床、現実の材質感と奥行きを参照する。壁面の十字架や宗教記号は使用しない。
Image4: 30秒プレビズ。CUT 01〜06の人物配置、画角、中央9:16セーフ、カメラ移動、カメラ高、アングル、終端構図を参照する。画像内の人物は演出確認用マネキンであり、完成映像ではImage1とImage2のキャラクターへ置き換える。
Image5: 立ち姿の灰色の兎のぬいぐるみ。現実的な体型、毛足、垂れ耳、暗い赤色のガラス製の瞳を参照する。
固定条件:
30秒、16:9横型、視聴者がその場にいるPOV視点。9:16へ中央トリムしても双子の顔、手、兎のぬいぐるみ、主要な演技が残るよう、関心事を常に中央へ集める。
実写とセル画アニメの明確なハイブリッド映像。部屋、家具、寝具、ランプ、窓、床、兎のぬいぐるみは実写フォトリアル。双子だけを高密度なセル画アニメとして描く。キャラクターの接地影、落ち影、環境光、ランプ光、遮蔽、色かぶりは実写背景と物理的に一致させる。
背景は全編で深い被写界深度。絞りF5.6〜F8相当の見え方とし、前景・中景・後景のレース、ベルベット、木目、金属、ガラス、カーテン、寝具、床、壁の質感と部屋の立体構造を明瞭に保つ。背景を雰囲気だけのボケにせず、実写空間として読める情報量を維持する。自然なレンズ描写、現実的なハイライトの減衰、微細なセンサーノイズ、実写らしい影と反射。
双子は瓜二つの美少女アニメキャラクター。Image1とImage2を最優先し、顔立ち、目つき、銀髪ツインテール、額と耳が見える髪型、黒いリボン、黒いゴスロリ衣装、黒い光輪を全カットで維持する。キャラクター配色は均一なベタ塗り、段階的な色調、強いセルシェーディング。主線は太さが安定した、ムラのない鮮明な黒線。2024年制作の日本ヴィンテージセルアニメ映画を思わせる精密な作画、高い描き込み、滑らかな動き。
双子の瞳は全編で澄んだグレー。白目は清潔な白。白いまつ毛、白いアイライン、白い上下まぶたの線、白い目の輪郭線、長い下まつ毛を維持する。目を赤くしない。瞳、白目、目の縁に充血や血管を一切出さない。
赤色は兎の暗赤色のガラス眼と、最後に双子の頬へ現れる上品な朱色の紅潮だけに限定する。頬の紅潮は目や鼻へ広げず、白い肌との色彩コントラストを明瞭にする。
カメラはPOVらしい距離感。全編にごく軽い手持ちの揺れと呼吸に同期する微細な上下動を入れるが、顔と背景の可読性を損なわない。各カットでは記載した主要カメラワークを優先し、不要な複合移動を加えない。
音楽なし。環境音もできるだけ入れない。日本語の女性ボイス、近距離で静か、囁きに近い自然な発声。台詞は指定どおりに発話し、字幕や画面文字は出さない。
全体設計:
前半15秒は静かな執着と親密さを積み上げる。後半は「裏切ったら……」の余韻から静かな笑いへ移り、3秒だけ照明とPOVの後退によるホラーへ反転し、最後の9秒で緊張を甘く裏返す。結末の主役は、澄んだグレーの瞳、青白い肌、朱色に染まる頬、蠱惑的かつ恍惚とした微笑。ホラーを引きずらず、怖さより可愛さと危うい魅力を最終印象にする。
[00:00-00:05]
動作・演出: 実写のゴシック寝室。妹が天蓋ベッドに腰掛け、ベッド上で立ち姿を保つ灰色の兎のぬいぐるみの頭を、指先でゆっくり撫でる。ランプの実写炎がわずかに揺らめき、妹の白い髪と黒い衣装へ硬いセル影を落とす。妹は撫でる手を止めず、最後に視線だけをカメラへ向ける。
カメラ: MCU、アイレベル0度。妹の上半身から顔のCUへ、0.8mのスロードリーイン。中央9:16セーフ内に顔、手、兎を維持する。
音声: 衣擦れ、毛足を撫でる小さな音、ランプの微かな燃焼音、遠い室内音。妹「コウテイペンギンっていいよね。」
[00:05-00:10]
動作・演出: 姉が妹の背後から静かに入り、妹の首から上胸部へ執着するように両腕を回して抱きつく。締め付ける暴力ではなく、逃がさない親密さ。二人の黒い光輪が不規則に弱く明滅する。二人は澄んだグレーの瞳でカメラをまっすぐ見つめる。
カメラ: MS、ハイアングル35度。頭上から0.6mのゆっくりしたクレーンダウン。終端で二人の顔を中央9:16セーフへ重ねる。
音声: ベッドの沈む音、布の擦れる音、低い室内残響。姉「アデリーもいいわよ。」
[00:10-00:15]
動作・演出: 二人の頬と髪が触れるほど顔を寄せる。表情は無機質な静けさから、口元だけがわずかに歪む危うい微笑へ移る。瞳はグレーのまま澄んでおり、背景には実写の天蓋、ランプ、レース、木製家具が鮮明に残る。
カメラ: CU、ローアングル15度。左の顔から右の顔へ20度だけゆっくりパンし、二人の間へ戻ってセンターで止める。
音声: 部屋の環境音を少し絞り、浅い呼吸を近くで聞かせる。妹「イワトビも……すき」
[00:15-00:18]
動作・演出: 二人は密着したまま、視線をカメラから外さない。妹の口元だけが小さくほどけ、姉はその笑いを共有するように僅かに頬を寄せる。目には異常を起こさず、静止に近い演技で不穏な間を作る。
カメラ: タイトCU、アイレベル0度。前カットのセンター位置から継続し、0.15mだけ極めて遅いマイクロプッシュイン。
音声: 0.4秒の静かな間のあと、妹の小さな含み笑い「フフッ」。BGMなし。
[00:18-00:21]
動作・演出: 一瞬だけホラーへ反転する。枕元のランプが乾いた音と同時に消え、冷たい窓光だけが実写の部屋の輪郭、レース、木材、二人のシルエットを薄く残す。二つの黒い光輪が闇の上で短く明滅し、兎の暗赤色のガラス眼が反射光を一点だけ返す。双子は動かず、目を赤くしない。最後の0.3秒でランプが再点灯し始める。
カメラ: MCU、ダッチアングル4度。POVが反射的に0.25mだけ鋭く後退するワンアクション。後退後は揺れを止め、暗闇を1拍だけ保持する。
音声: ランプが消える乾いたクリック、短い低周波の衝撃音、直後に無音へ近づく。叫び声、金切り声、ホラー音楽は入れない。
[00:21-00:30]
動作・演出: 暖かなランプ光が滑らかに戻り、恐怖の硬さが解ける。二人はカメラへ少しだけ身を寄せる。妹が「なんて……ね」と甘く囁き、姉も同時に目を細める。二人の白い肌は青白く清潔なまま、頬だけが時間をかけて上品な朱色へ染まる。二人はカメラを受け入れるような蠱惑的かつ恍惚とした微笑を浮かべる。澄んだグレーの瞳には柔らかなランプのキャッチライト。赤い光や充血ではなく、白い肌と朱色の頬の対比で陶酔感を表現する。実写背景のレース、天蓋、木目、ランプ、兎の毛足は最後まで鮮明に見える。
カメラ: CU、アイレベル0度。前カットのダッチを水平へ戻してから、二人の顔へ0.35mの非常に遅いドリーイン。終端は二人の顔を左右対称に中央へ置き、頬と目、背景の実写質感が同時に読める距離で停止する。
音声: ランプが戻る小さな点灯音、寝具の僅かな沈み、近い呼吸。妹「なんて……ね」。語尾は蠱惑的で甘く、笑みを含ませる。声色と表情で表現する。
終了フレーム:
アイレベルの左右対称CU。二人が頬を寄せ、澄んだグレーの瞳でカメラを見つめ、白い肌と上品な朱色の頬を対比させた、蠱惑的かつ恍惚とした微笑。黒い光輪は安定して弱く輝く。実写のランプ、天蓋、レース、木目、兎の毛足を深い被写界深度で鮮明に残し、0.5秒静止保持する。
避けること:
目の充血、赤い白目、赤い虹彩、赤い瞳孔、目の血管、赤い目の縁、泣き腫らした目、目周辺への赤い照明。最終カット以外の頬の強い紅潮。頬の赤みが目や鼻へ広がる表現。
浅すぎる被写界深度、背景ボケ、クリーミーなボケ、顔以外の全面ぼかし、実写背景の情報消失、イラスト背景、アニメ背景、絵画化した家具、人工的な被写界深度処理。
キャラクターの実写化、背景や兎のアニメ化、セル画キャラクターと実写空間の影・光・接地の不一致。
Image1の構図の複製。壁の十字架、磔刑像、宗教記号。流血、傷、暴力、首を締める表現、過度にグロテスクな演出、長すぎる暗転。
双子の顔、髪型、衣装、体型、光輪の変形や入れ替わり。第三の人物、余分な腕や指、融合した身体、兎の増殖や変形。
字幕、台詞文字、ハート記号、ロゴ、透かし、UI、カメラ注釈を完成映像へ表示すること。
激しい手ブレ、過剰なローリング、急すぎるズーム、顔や手の見切れ、中央9:16セーフから主要演技が外れること。5b:T799,Create a fast, trippy 30-second dubstep music video using all the uploaded photographs
(PLEASE FIND ATTACHED ALL 14 IMAGES + AUDIO IN THREAD).
Use the uploaded 30-second audio as the only soundtrack. Make it a powerful dubstep edit: every movement, punch-in, transition and hard cut must hit the beat. Fully animate everything inside the photographs. Make the dancers perform, fabric and flames spin, painted eyes react, masks lunge, multiple arms ripple, the drummer strike his drum, sadhus dissolve into ash and black birds, motorcycles race aggressively, smoke creatures come alive, the underwater band actively play, boats move through the surreal landscape, and the flaming weapon create violent circles of fire. No photograph should remain static.Freely mix all the visuals in whatever order works best with the music. Let movement from one photograph naturally transform into the next: spinning fire becomes fabric, fabric becomes a painted eye, expanding arms become the drummer’s arms, ash birds become racing motorcycles, motorcycle dust becomes dancing smoke, smoke becomes underwater bubbles, bubbles become water and boats, and the boats transform into circles of fire.Keep constant movement throughout. Use rapid punch-ins, speed ramps, bass-impact flashes, hard rhythmic cuts, brief freeze-frames and fluid psychedelic transitions. Add a separate graphic layer flashing ॐ, the trishul, damru, third eye, crescent moon and traditional Hindu sacred symbols precisely on the beat. Let them pulse, rotate, distort and dissolve without covering the performers.Make it stylish, colourful, strange and extremely high-energy—an edgy, award-winning Incredible India advertisement edited like a dubstep music video. No dialogue, no typography, no slideshow, no long static shots and no complicated story. Finish with the strongest collision of fire, ash, birds and sacred symbols on the final bass hit.
#capcutseedance25 #capcutcpp5c:T104c,---------------
Duration: 20s | 16:9 | 24fps Duel from Romance of the Three Kingdoms, battle at Hulao Pass, late Han dynasty. REFERENCES
Image 1
— 呂布 LU BU
Image 2
— 関羽 GUAN YU
Image 3
— 張飛 ZHANG FEI Use them for IDENTITY ONLY: face, beard, hair, headwear, armour, cloak colour, build, proportions. IGNORE their white studio background, flat studio lighting and standing pose — none of it appears here. Every shot is outdoors, and all three men are mounted and fighting. STYLE (all beats) A dusty plain before a huge stone mountain pass at dusk. Two vast armies on either side, banners rippling, dust in the air. Overcast sky, low copper sun breaking through. Palette: iron grey, ochre dust, blood red, bronze. Desaturated and earthy. 50mm, shallow focus, heavy haze, film grain, high contrast, hard rim light through the dust. Photorealistic historical epic, sweat and grime, worn scratched armour. No fantasy glow, no magic. WEAPONS — no reference exists. Build from this, keep identical throughout. Each is 1.5x its owner's height. Never swap them. 呂布: ji halberd — straight spear point, one crescent side blade below it on one side, long dark shaft, bronze fittings. 関羽: guandao — one heavy broad curved blade with a hook on its back edge, joined to a long dark shaft by a bronze dragon head. 張飛: serpent spear — slender steel blade forged in an S-shaped wave, on a plain dark shaft bound with cord. HORSES — no reference exists. Anatomically correct, four legs, correct joints. 呂布: large red-chestnut stallion, black mane, bronze fittings, red tassels. 関羽: tall dark bay, black mane and lower legs, plain leather tack. 張飛: heavy solid black, thick neck, shaggy fetlocks, battered iron tack. BEAT 1 (0-4s) 呂布 — clean-shaven, gold headdress with two pheasant feathers, red cloak, halberd, red-chestnut horse — sits alone between the two armies, halberd across his saddle. The horse stamps and turns a half step. He raises the halberd and levels it at the enemy line. Camera: WS from low, slow push-in. BEAT 2 (4-8s) 張飛 — bristling beard, black iron armour, dark cloak, serpent spear, black horse — bursts from the line at a gallop, roaring, and drives the spear at 呂布. 呂布 turns the halberd across his body and knocks it aside. Sparks and dust. The horses wheel past each other. Camera: tracking alongside at horse height, moving with the charge. BEAT 3 (8-12s) 関羽 — long smooth black beard, ruddy face, dark green robe over armour, green headwrap, guandao, dark bay horse — rides in from the far side, rises in the stirrups and cuts down overhead. 呂布 catches it on the haft. Both strain, blades locked, horses shoulder to shoulder. Camera: MCU, orbiting once around the locked weapons. BEAT 4 (12-16s) All three circle in churning dust, weapons striking and turning each other aside, horses rearing. 呂布 in red and gold, 関羽 in green, 張飛 in black iron. Cut tight to 呂布's eyes under the headdress — narrowing, calculating, uncertain for the first time. Camera: handheld, close, drifting between them. BEAT 5 (16-20s) 呂布 wrenches his horse around, sweeps the halberd once to drive them back, and gallops for the pass, red cloak and feathers streaming. 関羽 and 張飛 pull up side by side, chests heaving, weapons still raised, watching him go against the low copper sun. Camera: WS, pulling out to reveal both armies. AUDIO War drums and massed shouting throughout, rising through beats 2-4. Hooves on dry earth, snorting horses, steel ringing on steel, creaking leather, wind and dust. A deep roar from 張飛 at 4s. Near-silence for a breath at 12s on 呂布's eyes, then the clash returns. No dialogue. AVOID white studio background, flat studio lighting, seamless backdrop, neutral standing poses, facial hair on 呂布, on-screen text, subtitles, watermarks, modern objects, fantasy glow, magic, wire-work floating, clean unworn costumes, faces/armour/weapons swapping between beats, weapons changing shape or length, horses with deformed or extra legs, extra or missing limbs, more than three warriors, mixed camera moves, jitter, flicker, ghosting.5d:Tdf1,Lock the original illustration completely. Preserve the exact art style, brushwork, textures, material rendering, color palette, creature proportions, facial expressions, lighting, composition, and every character design exactly as shown. Never change the visual style. Every creature must remain identical in every shot. No text, watermark, logo, or signature. One continuous story. Motion always progresses forward. No creature moves backward. Use cinematic pacing with hard cuts, speed ramps, bullet time, dynamic tracking, volumetric water, realistic physics, motion blur, subtle film grain, dramatic lighting, lens droplets, cinematic depth of field, bold camer
Lock the original illustration completely. Preserve the exact art style, brushwork, textures, material rendering, color palette, creature proportions, facial expressions, lighting, composition, and every character design exactly as shown. Never change the visual style. Every creature must remain identical in every shot. No text, watermark, logo, or signature. One continuous story. Motion always progresses forward. No creature moves backward. Use cinematic pacing with hard cuts, speed ramps, bullet time, dynamic tracking, volumetric water, realistic physics, motion blur, subtle film grain, dramatic lighting, lens droplets, cinematic depth of field, bold camera language, emotional facial animation, and seamless transitions.
0-3s
Without warning, the peaceful group suddenly notices every bird exploding into the sky. Tiny ripples appear beneath their feet. A deep rumble grows louder. Extreme close-up of frightened eyes. Fast push-in toward the rabbit. Hard cut to a gigantic wall of water rising behind them. Lightning-fast whip pan reveals the impossible tsunami racing toward the camera. Deep bass hit.
3-6s
The wave crashes over everyone. Water engulfs the frame. Camera tumbles underwater with bubbles, debris, flowers, leaves and sunlight cutting through the current. Every creature is separated naturally by the force. Slow-motion for one second before time accelerates again.
6-10s
Rapid cross-cut survival sequence. The rabbit paddles frantically with long ears floating behind. The lion powers through the waves with strong strokes. The rhino kicks heavily while struggling to stay above water. The giant blue triceratops swims like a massive whale. Frogs leap from floating debris. Tiny creatures cling to drifting branches. Cats desperately paddle while trying not to sink. Dynamic handheld water-level camera alternating with overhead drone shots.
10-15s
The current becomes more violent. Floating logs collide. The turtle spins helplessly in a whirlpool before escaping. The sheep nearly slips beneath the surface until the lion nudges it upward. The black cat loses its grip and disappears beneath a wave. Silence for half a second except muffled underwater sound.
15-20s
The gray kitten dives underwater searching. Cinematic underwater tracking follows the kitten weaving between drifting plants and bubbles. A frightened turtle struggles below. The kitten reaches it, grabs its shell with tiny paws and guides it upward toward sunlight breaking through the surface.
20-25s
Both burst above the water. The exhausted kitten climbs onto the turtle's shell. The turtle begins swimming steadily through the endless flood while the kitten balances carefully, soaked but relieved. Around them every animal continues fighting the current in its own unique way. Long aerial tracking shot follows their path.
25-28s
The current slowly weakens. One by one the surviving animals regroup behind the turtle. The rabbit swims beside it. Frogs hop onto floating wood. The lion keeps everyone together. The enormous blue triceratops shields the smaller creatures from remaining waves. Warm sunlight finally pierces the storm clouds.
28-30s
A majestic vertical drone shot rises high into the sky. The flood forms beautiful winding rivers across the landscape. At the center, the tiny kitten sits safely on the turtle's shell leading the group forward while every surviving creature follows behind. The music swells into an emotional orchestral finale. Hold the final frame like an animated movie poster as golden light reflects across the water.
SUBJECTS:
セレスティア(Celestia)、闇の魔導師 @ image1。銀白色から淡いラベンダーへとグラデーションする長い髪、幻想的に輝く紫色の瞳を持つ優雅な闇の魔法使い。発光する紫のクリスタル装飾が施されたゴシック調の黒いドレスをまとい、宙に浮かぶ紫のクリスタルコアを頂いた洗練された黒い魔法の杖を携える。
ENVIRONMENT:
雲の上に浮かぶ壮大な魔法都市。浮遊島、石造りの空中橋、魔法使いの塔、巨大なクリスタル尖塔、空中に浮かぶ書店、魔法カフェ、飛行船、浮遊庭園、上向きに流れる幻想的な滝、浮遊ランタン、魔法の樹々が広がる幻想世界。
CAMERA(全体):
全編を通してスピード感のあるシネマティックFPV視点。各カットでは「1つの主要カメラワーク」のみを採用し、演出を整理。カット間はホイップパンやモーションブラー・マッチカットで自然につなぎ、ハードカットは使用しない。
STYLE:
最高品質のアニメシネマティックスタイル。超高品質イラスト、鮮やかなファンタジーライティング、美しいセルシェーディング、力強いアニメ線画、絵画的ハイライト、幻想的なブルーム、ボリューメトリックゴッドレイ、魔法のレンズフレア、AAA級ファンタジーアニメのオープニング品質。
COLOR LOGIC:
漆黒、ミッドナイトブルー、バイオレット、ラベンダー、輝くマゼンタ、純白のハイライト、淡いシアンの環境光、暖かなゴールドのランタン光、エメラルドグリーンの植物。魔法エフェクトは深い紫を基調とする。
━━━━━━━━━━━━━━━━━━━━━━
TIMELINE(合計30秒)
━━━━━━━━━━━━━━━━━━━━━━
0:00-0:05【静止 → 爆発的始動】
MAIN FX:巨大な魔法陣1つのみ
セレスティアは時計塔の頂上で静かに立ち、マントと髪が風になびく。魔法の粒子が漂う中、杖を掲げると足元に巨大な魔法陣が展開する。
CAMERA:
ゆっくりとしたオービット → 急激なスピードランプ
TRANSITION:
モーションブラー・マッチカットで屋根上チェイスへ
SFX:
魔力の低い唸り、クリスタル共鳴、魔力爆発、突風
━━━━━━━━━━━━━━━━━━━━━━
0:05-0:10【疾走】
MAIN FX:
足元に生成される発光魔法足場のみ
屋根の上を高速で駆け抜け、一歩踏み出すたびに足元へ魔法の足場が出現し、それを蹴って次々と跳躍する。他の魔法演出は重ねない。
CAMERA:
FPV追跡、足元中心のローアングルトラッキング
TRANSITION:
大きく跳躍しながら上方向へホイップパン
SFX:
魔法の足音、魔力バースト、風切り音
━━━━━━━━━━━━━━━━━━━━━━
0:10-0:16【静止 → ポータル連鎖】
MAIN FX:
転移ポータル
空中で約0.4秒だけスローモーション。周囲が静まり粒子だけがゆっくり漂う。その直後、高速で複数の転移ポータルを連続通過する。
CAMERA:
スローモーション静止 → ポータル回転に合わせた高速オービット
TRANSITION:
ポータルの白い閃光で街への急降下へ
SFX:
(0.4秒ほぼ無音)
→ ポータル共鳴、クリスタルベル、空間歪曲音
━━━━━━━━━━━━━━━━━━━━━━
0:16-0:21【急降下魔法】
MAIN FX:
街全体を照らす巨大ルーン1つ
ランタンが並ぶ街路へ急降下しながら巨大魔法を発動。街全体が紫色の光で照らされる。蝶やリボンは背景演出程度に留める。
CAMERA:
急降下ショット+クローズアップ+ワイド俯瞰を交互に使用
TRANSITION:
スピードランプでラストスプリントへ
SFX:
多層魔法詠唱、クリスタルベル、魔法爆発
━━━━━━━━━━━━━━━━━━━━━━
0:21-0:27【クライマックス】
MAIN FX:
巨大衝撃波
最後の全力疾走。クリスタル尖塔へ跳躍し空中で一回転。着地と同時に杖を地面へ突き立て、都市全体へ巨大な魔力衝撃波が広がる。この瞬間のみ演出を最大化する。
CAMERA:
空中回転中はオービット → 着地後は都市全景へ大きく引く
TRANSITION:
ワイドショットを維持しながら衝撃波の光が徐々に消えていく
SFX:
壮大な魔法クレッシェンド、クリスタル噴出音、オーケストラのインパクト
━━━━━━━━━━━━━━━━━━━━━━
0:27-0:30【余韻】
MAIN FX:
残光のみ
セレスティアは尖塔の頂上に静かに立ち、衝撃波から舞い降りる光の粒子がゆっくり降り注ぐ。タイトルやロゴを配置できる余白として使える締め。
CAMERA:
ゆっくりとしたプッシュイン → 静止
SFX:
消えていくクリスタル共鳴、静かな魔力の余韻61:T98b,A 30 second film from ONE completely locked, static camera position that never moves, never pans, never zooms, never cuts. The same view of the same patch of ground for one thousand years, time passing at enormous speed. In the foreground of every single frame sits a large mossy grey boulder. It never moves. It is present and recognisable in every frame of the film from first to last. [0:00-0:03] Ancient temperate forest. Huge old trees, deep green undergrowth, mist between the trunks, soft grey daylight. [0:03-0:06] Trees fall and are dragged away. A clearing opens. Stumps, churned mud, a first thin column of smoke. [0:06-0:09] A single timber and thatch hut rises beside the boulder. A small vegetable plot. Chickens. [0:09-0:12] A village grows around it. More huts, fences, a dirt track, wood smoke, sheep in the field beyond. [0:12-0:15] A stone wall is built, incorporating the boulder into its base. A church tower rises behind. Snow falls and melts. Seasons flicker past. [0:15-0:18] Fire. The village burns, timbers collapse, smoke blackens the sky. The wall and the boulder survive. [0:18-0:21] Rebuilding in brick. A cobbled road. A chimney, then many chimneys. Industrial smoke, grey air, rain. [0:21-0:24] The town becomes a city. Concrete rises. The old wall is demolished. The boulder sits in a small paved square, surrounded by buildings. [0:24-0:26] Glass towers. Traffic light streaking past at night. The boulder is fenced off, a small green plaque beside it. [0:26-0:28] Abandonment. Windows dark, glass cracked, weeds through the paving, ivy climbing the towers. Silence. [0:28-0:30] The towers have collapsed and gone. Young forest has grown over everything. Mist between the trunks. The frame is now identical to the opening frame, the same boulder in the same position under the same soft grey light. Camera: absolutely locked and static for all 30 seconds. No movement of any kind. No cuts, no dissolves, no transitions — one continuous transforming shot. Style: photoreal, cinematic, natural overcast light throughout, muted natural colour, seasons and weather flickering rapidly past. Audio: diegetic only, evolving with the era — forest birds, axes, village life, church bell, fire, industrial machinery, city traffic, then silence, then birds again. NO MUSIC, no score, no soundtrack, no instruments. No text, no captions, no titles, no watermarks, no camera movement, no cuts. created with @higgsfield62:T1383,Broadcast footage from a real 2000s–2010s Japanese sports entertainment program. Realistic TV broadcast footage reminiscent of a major Japanese obstacle course competition show. Full-scale sports coverage combining ENG cameras, crane cameras, wire cameras, and telephoto cameras. Natural handheld camera shake, zoom, pan, and tracking shots. Slightly soft footage characteristic of early HD broadcasts, natural broadcast noise, TV compression feel, realistic color reproduction. The venue features large lighting, spectator seating, announcer's booth, and the enthusiastic cheers of the heated audience. A tense atmosphere of anticipation and the challenger's seriousness permeate the entire competition venue. Fix the main subject completely to @[Image 1](image_1). Do not change the r
Use the provided image as the strict visual reference for the warrior’s face, blonde hair, armor, fur jacket, leather boots, sword, body proportions, ruined battlefield, burning fortress, smoke, embers, and dark cinematic atmosphere. Preserve her identity and outfit throughout the entire sequence.
Create a 20-second impossible high-fashion warrior photoshoot, blending sensual confidence, brutal fantasy action, luxury editorial posing, explosive VFX, and aggressive cinematic camera work. The result must feel like a viral fashion campaign filmed in the middle of a medieval war.
0–4 seconds — Iconic opening pose
Begin with the exact composition of the reference image. The warrior stands on shattered stone, sword lowered beside her, staring directly into the lens. Her hair and fur move violently in the wind while embers spiral around her. The camera rushes toward her with a low-angle dolly-in. A powerful white flash freezes her pose, followed by a horizontal blue lens flare crossing her eyes and armor.
4–8 seconds — Fashion meets combat
She walks toward the camera with controlled, sensual confidence as explosions erupt behind her. Use rapid editorial cuts: close-up of her eyes, lips, hand tightening around the sword, fur texture, armored waist, boots crushing burning debris. She suddenly pivots and blocks an incoming blade from an unseen attacker. Sparks explode across the lens as the camera whip-pans around her.
8–12 seconds — Impossible action photoshoot
She launches into a fast spinning sword movement, cutting through smoke and shattered banners. Freeze her mid-motion in a dramatic fashion pose as dozens of camera flashes fire from impossible angles. Add elegant motion echoes, metallic silhouette trails, floating reflective shards, streaking embers, and anamorphic lens flares. Each flash captures a different powerful pose before snapping back into live action.
12–16 seconds — Sensual hero moment
The action suddenly slows. She plants the sword into the ground, steps close to the lens, and leans slightly forward with a fierce, provocative gaze. Wind sweeps her hair across her face while warm firelight shapes her silhouette. The camera performs a tight orbit around her body and rises into a close-up. A strong golden lens flare blooms behind her shoulder as ash drifts through the frame.
16–20 seconds — Epic viral finale
A massive explosion destroys part of the fortress behind her. Without looking back, she pulls the sword free and turns sharply toward the camera. Use an extreme speed ramp as she swings the blade directly past the lens, creating a seamless transition into three rapid flash-frozen editorial poses: sword over her shoulder, blade pointed downward, then a dominant frontal stance.
Finish on a slow-motion hero frame with the warrior standing perfectly centered, smoke and fire behind her, embers suspended in the air, her hair flowing, armor glowing along the edges, and a huge cinematic lens flare crossing the screen.
Ultra-premium fantasy fashion editorial, sensual but powerful warrior energy, impossible battlefield photoshoot, luxury magazine attitude, brutal sword choreography, low-angle camera, macro beauty shots, whip-pans, fast orbit movement, snap zooms, speed ramps, flash photography, freeze frames, motion echoes, metallic VFX, realistic fire and debris, dramatic anamorphic lens flares, realistic hair and fabric physics, no dialogue, no text, no subtitles, no identity drift, no wardrobe change, no facial deformation, no extra limbs.67:Tef4,@Image1 is the
One character reference, rebuilt as a fictional person
参照キャラクター画像を1枚使用する。
参照画像の髪型、体型、衣装シルエット、表情の方向性、雰囲気を参考にする。
実写MV風の架空人物として自然に再構成する。
実在の俳優、モデル、有名人、特定の人物には似せない。
参照画像の衣装、色、アクセサリー、世界観をもとに、背景・VFX・色・記号・装飾をランダム生成する。
背景は固定せず、参照画像に合う実写MV風の場所をランダムに選ぶ。
屋上、路地、ステージ裏、廃ビル、夜の街、ネオン空間、グラフィティ空間、ライブ会場裏など。
全体はグラフィティ系実写MV風に統一する。
60fps。
118 BPM基準。
高品質な実写MV。
リアルな人物、肌、髪、布、アクセサリー、靴。
自然な照明と現実のカメラ質感。
背景には実写と馴染むグラフィックモーションVFXを重ねる。
ダンスは、男性にも女性にも似合う自然なストリート系MVダンス。
難しい技を連発しない。
人間が自然に踊れる範囲の動きにする。
脚を曲げ続けず、自然な立ち姿勢を保つ。
膝はビートのアクセントで軽く沈む程度にする。
重心移動、踏み替え、軽いターン、肩のリズム、手先の動きを中心にする。
動きはキレがあるが、ロボットのように硬くしない。
身体全体が音に乗って自然につながるようにする。
単純な横移動を主役にしない。
前後の奥行き、身体の角度変化、軽い踏み替え、自然なターンで見せる。
腕は振り回さず、身体に近い位置でスタイリッシュに動かす。
手首や指先は自然に使う。
過剰にセクシーにしない。
アイドルダンスにも、バレエにも、ロボットダンスにも寄せない。
背景とVFXは常に軽く動く。
参照画像の世界観に合う抽象記号、短い文字断片、模様、アイコン、スプレー跡、光の線、ノイズ、図形、装飾をランダムに配置する。
文字や記号の意味は固定しない。
VFXは人物の後ろで控えめに反応し、顔や身体を隠さない。
VFXは身体の動きに自然に反応する。
足の接地で小さな粒子や波紋が出る。
肩や手の動きに合わせて背景の線や光が軽く揺れる。
衣装の差し色から小さなスパークや色残像が出る。
布やアクセサリーの揺れに合わせて短い光の残像が出る。
演出は派手すぎず、実写MVに馴染ませる。
カメラは実写MVの自然なカメラワーク。
開始直後から中距離の全身ショット。
人物は映像開始時点ですでに踊っている。
カメラは完全停止しない。
軽い押し込み、軽い引き、斜め前からの小さな回り込み、短い足元カットを使う。
カメラを動かしすぎない。
ダンスのライン、衣装、表情が読めるように撮影する。
0.0s - 3.5s
開始直後から自然に踊っている。
中距離の全身ショット。
片足を軽く前へ踏み込み、ビートに合わせて上半身が自然に乗る。
肩を小さくヒットさせ、手首を軽く外へ流す。
足元は小さな踏み替えとつま先タップ。
膝は軽く使うだけで、曲げ続けない。
表情は参照画像の雰囲気を保つ。
背景VFXは軽く跳ね、足元に小さな粒子が出る。
3.5s - 7.0s
カメラが斜め前から少しだけ押し込む。
人物はその場を中心に、前後の踏み替えと軽いターンで踊る。
腕は身体の近くでコンパクトに動かす。
片手を衣装ディテールの近くへ持っていき、軽く離す。
肩、胸、手首の動きが滑らかにつながる。
脚のラインが自然に伸びる瞬間を作る。
背景のランダムなグラフィックが音に合わせて揺れる。
7.0s - 10.5s
上半身を少し強調する。
短い視線、軽い首の角度変化、手先のフリックを入れる。
足元は大きく移動せず、踏み替えと小さなクロスステップでリズムを取る。
身体を少し斜めに向けてから正面へ戻す。
ターンは小さく自然にする。
VFXは衣装カラーに合わせて、光の線や粒子が控えめに反応する。
10.5s - 13.5s
動きを少しだけ強くする。
前後のステップ、肩ヒット、手首スナップ、軽い胸ヒットをつなげる。
一つ一つの動きは独立させず、身体全体の流れとして自然につなぐ。
膝は一瞬だけ沈め、すぐ自然に伸ばす。
背景グラフィックが少し活発になり、奥行き方向へ流れる。
13.5s - 15.0s
次の動きへつながる準備。
人物は前後のクロスステップと小さなターンを続ける。
最後に止まらない。
決めポーズにしない。
身体がまだ動いていて、次のステップへ入る途中で終わる。
背景VFXも止めず、奥へ流れるように動かす。
全体印象:
参照画像ベースのカラフルなグラフィティ系実写MVダンス。
自然に人間が踊っているような、滑らかで見栄えの良い振付。
男性にも女性にも似合う。
キレはあるが、硬すぎない。
脚を曲げ続けず、自然な立ち姿勢と軽い沈みを交互に見せる。
横移動を多用せず、前後の奥行き、角度変化、踏み替え、自然なターンで見せる。
開始直後からダンスで入り、次の動きへ繋がる途中で終わる。
テキストオーバーレイなし。
字幕なし。
ロゴなし。78:Tf80,Authentic live-action cinematic footage captured as a continuous one-shot sequence on Arri Alexa 65 with Panavision Ultra Vista anamorphic lenses on real Kodak Vision3 35mm film. Ultra-photorealistic Hollywood adventure film aesthetic with natural film grain, cinematic depth of field, realistic handheld camera inertia, dynamic tracking shots, physically accurate motion blur, volumetric golden-hour sunlight, atmospheric jungle mist, practical lighting, realistic water simulation, authentic environmental interaction, and absolutely no CGI or game-engine appearance.
A powerful Amazonian tribal warrior with weathered bronze skin, long braided black hair decorated with jungle feathers, athletic muscular build, worn leather tribal clothing, and intense recognizable eyes sprints through a dense rainforest while the camera performs an energetic low-angle tracking shot weaving between giant roots, rocks, vines, and tropical vegetation. He bursts out onto the edge of a colossal waterfall and leaps into the air as the camera dives beside him. During freefall his body organically transforms into a gigantic Amazon river crocodile with realistic skeletal restructuring, growing scales, extending jaws, muscular limbs, and a powerful tail before crashing into the river below.
The camera follows underwater in one continuous movement as the crocodile powers through crystal-clear water, scattering schools of fish and passing submerged trees and ancient stone ruins. While swimming at high speed, the crocodile seamlessly transforms into a massive Amazon giant catfish with smooth biological morphing, retracting limbs, sleek skin, broad fins, powerful tail propulsion, and continuous identity preservation. The giant catfish accelerates through flooded jungle ruins before its body thickens and organically transforms into a colossal emerald anaconda with shimmering green scales, enormous muscular coils, and realistic serpentine movement as it glides through submerged temples and rises onto the rainforest floor.
The colossal anaconda rapidly slithers through dense jungle while the camera performs fast ground-level tracking through mud, roots, and foliage. As it launches over a fallen tree, it seamlessly transforms into a powerful black jaguar with sleek black fur, muscular feline anatomy, realistic shoulder movement, explosive sprinting speed, and intense recognizable eyes. The jaguar races through the rainforest before reaching a towering cliff where it leaps into open air. During the jump, its body organically transforms into a majestic giant harpy eagle as black fur becomes feathers, forelegs expand into enormous wings, hind legs become massive talons, and the eagle catches an air current before soaring high above the endless Amazon canopy.
The camera transitions into an FPV-style aerial pursuit as the harpy eagle flies between clouds toward an ancient mountain temple surrounded by waterfalls. The eagle folds its wings into a dramatic dive toward the summit. During the dive, feathers organically become skin, wings become human arms, talons become feet, and the original Amazonian warrior fully reforms with perfect facial identity, braided hair, and worn tribal clothing. He lands powerfully on the temple platform, rises slowly, and looks directly into the camera while the camera performs a slow cinematic push-in. Behind him, endless rainforest, waterfalls, drifting mist, and golden sunlight create an epic final hero shot before a slow fade to black.
Maintain one continuous uninterrupted shot, seamless transitions, consistent character identity, realistic anatomy, believable body weight, physically accurate transformations, natural cloth s
1Write it in one fixed order. Subject, then action, then location, then camera work, then lighting, then picture texture. "A person walking" gets you a stock loop; the same subject with a camera move and a light direction attached gets you a shot. Creators who have spent real credits on 2.5 converge on this order.
2Thirty seconds is written as stages, not as longer sentences. The official guide splits a long clip into [STAGE ONE] / [STAGE TWO] / [STAGE THREE], each with an opening state, one main event, and an end state the next stage inherits. Padding a 10-second prompt with adjectives does not make a 30-second clip; splitting the event into inherited states does.
3One main action per time segment, described exactly once. Whether you use stages or plain 0-5s / 5-10s ranges, give each segment a single main movement. Two things go wrong otherwise: three actions crammed into five seconds all get compressed, and the guide is explicit that describing the same action twice degrades it.
4Give every reference asset a job and an exclusion. Not "here are my images" but "@image1 supplies the armour design and colour breakdown. Do not take the background." You can attach up to 50 assets; the more you attach without assigning roles, the faster stability drops.
5Name your subjects and reuse the name. Write <SOLDIER>, <SHARI>, <WOMAN> once and refer to that token every time after. Over 30 seconds, "she" and "the man" drift onto the wrong body — a named token is the cheapest identity lock there is.
6Write the sound as a cue sheet, not as a mood. Audio comes out of the same pass as the picture, so list what you want: ambience, the specific foley, dialogue with the language named, and what should be absent. Several prompts in this library end with a full sound list, and it is a large part of why they look finished.
7Ask for the imperfections you want. Handheld shake, autofocus that settles late, exposure breathing, framing that clips a face. Realism in this model is a list of specified flaws — leave them out and you get a clean commercial no matter how many times you write "candid".
8End with the negative list, and keep parameters out of it. Close with what must not appear: no subtitles, no watermark, no extra limbs, and — for reference-driven jobs — never render the reference sheet itself or duplicate the subject. Resolution, duration and aspect ratio are generation settings, not prompt text.
Prompt formula
Subject + action + location + camera work + lighting + picture texture + sound. Past about ten seconds, wrap that in timed stages: each stage gets one main action and an end state the next stage inherits, then close with the negative list.
Example: <BARISTA> pulls the last shot of the night in a closing café and sets the cup on the pass. Rain on the window, chairs already stacked, one warm pendant lamp over the machine and the rest of the room in shadow. The camera holds a low medium on the portafilter, then rises to a profile as the cup lands. Photoreal, warm tungsten against cold window light, fine grain. Sound: the grinder winding down, steam hiss, rain, one mug on wood, no music. Do not add customers, subtitles or a watermark.
Frequently asked questions
What is Seedance 2.5?
Seedance 2.5 is ByteDance's video generation model, unveiled on 23 June 2026 at the Volcano Engine FORCE conference. On the shipped API it generates 4 to 30 seconds in a single pass (or duration -1 to let the model pick the length), takes up to 50 multimodal reference assets in one job (up to 30 images, up to 10 video clips and 10 audio clips), adds video editing and video extension as task types so a product or background can be swapped without re-rolling the whole clip, and generates audio in the same latent space as the picture. Output is 480p or 720p — the native 4K at 10-bit colour announced at FORCE is not something the API exposes. ByteDance also shipped an official prompting guide for it, which we have walked through in English.
Are these prompts free to use?
Yes. Every prompt on this page is free to copy and run anywhere you have access to Seedance 2.5 — no login, no payment.
Who wrote these prompts?
Two kinds, and they are labelled differently. Prompts the creator published are kept in the language they wrote them in, credited, and linked to the original post — copyright stays with the author. For clips whose creator never published a prompt, we reverse-engineer one from the finished video and mark it "reconstructed": it describes the output, so it cannot recover negative constraints, exact dialogue or the reference-image workflow. It is a writing reference, not the creator's prompt. If you are an author and want an entry removed or a credit corrected, email us.
Where can I run Seedance 2.5 today?
On apimodels.app, as the model seedance-2.5, since 8 August 2026 — 480p $0.134/s, 720p $0.300/s, charged only on success. That is a change: probing our production Volcengine Ark account on 4 August 2026, every Seedance 2.5 model id still returned InvalidEndpointOrModel.NotFound, and we said so rather than sell a page that errors on generate. One key covers it alongside Seedance 2.0 (still the one to use if you need 1080p, from $0.092/s), MiniMax H3, Kling, VEO and Grok. The creators in this library ran these prompts in ByteDance's own Dreamina (Jimeng) and in Higgsfield.
How is writing for Seedance 2.5 different from Seedance 2.0?
The craft is the same; the structure is not. At 4-15 seconds you can write one paragraph for one action. At 30 seconds you have to split the clip into stages, give each stage one main action and an end state the next stage inherits, and lock identity with named subject tokens instead of pronouns — otherwise the model spreads the action evenly and identity drifts across the back half. The reference workflow also scales: 2.0 takes up to 9 images plus a reference video and audio, 2.5 takes up to 50 assets, which makes assigning each one an explicit job and exclusion much more important.
How long should a Seedance 2.5 prompt be?
As long as the clip needs and no longer. The prompts in this library run from about 1,700 to 4,600 characters, and the longest ones are the 30-second pieces — not because they are wordier, but because a 30-second clip has more states to pin down. If your prompt is long because it describes the same action three ways, cut it; the official guide is explicit that repeating an action degrades it.
Run these prompts on apimodels.app
Seedance 2.5 is live here alongside 85+ other image, video, audio and language models — one API, one key, pay as you go, $1 free on sign-up.
The prompts we reconstructed are published in full under MIT; author-written prompts are indexed with credit and a link to the original post. Open an issue to suggest more.