Curated Prompt Vault
Xianxia Sisters Stoic Challenge Scene
Cinematic realism, pure ancient Chinese fantasy aesthetics, restrained yet comedic performances, observational cinematography, realistic micro-expressions, bod…
Goku Prompt Hub
Curated Prompt Vault
Cinematic realism, pure ancient Chinese fantasy aesthetics, restrained yet comedic performances, observational cinematography, realistic micro-expressions, bod…
Video Preview
Cinematic realism, pure ancient Chinese fantasy aesthetics, restrained yet comedic performances, observational cinematography, realistic micro-expressions, body weight, silk textures, delicate film grain, volumetric atmospheric depth, realistic 3D camera movement, and a reference world that continuously exists from the first frame to the last. Control Protocol—Three tracks must operate simultaneously, but none can usurp narrative control. The protagonist's storyline track has 100% control. The environment's lifeline track has continuous movement rights but 0% control. The background character track has independent life rights but 0% interaction rights with the protagonist. The environment cannot create, trigger, explain, interrupt, resolve, change, emphasize, or obstruct any character's storyline. Wind, clouds, fog, light, reflections, buildings, distant characters, vegetation, and water cannot react differently to dialogue, eye contact, comedic moments, or the protagonist's actions. The camera's movement is determined solely by the protagonist's storyline and character movement. However, at the same time: all non-protagonist areas must never be frozen. All currently uploaded reference images are understood as the DNA of the world, not frozen pixels or a static background that needs protection. Before the actual generation, the topographical logic, architectural vocabulary, scale relationships, materials, weather, cloud systems, vegetation, reflective surfaces, traffic routes, main light directions, foreground-midground-background relationships, and off-screen spatial extensions within the reference images are comprehensively understood. Then, a realistically connected 3D location is reconstructed. The identity and logic of the reference world are preserved, but the spatial arrangement and camera entry direction are allowed to be rearranged, without mechanically copying the original 2D composition of any reference image. Throughout the 10 seconds, each approximately 2-second time window requires multiple clearly visible non-protagonist movement sources at different spatial levels. These movements cannot start simultaneously, nor can they move at the same speed. Extreme distance: A huge, plausible cloud or air layer in the current world, which has been slowly and steadily migrating through the grand distant structure since before the scene begins. Deep midground: 4–6 small distant environmental characters are arranged as independent extras. One of them walks continuously for several seconds along a realistic passageway in the distance. Another person ascends or descends along existing steps or paths. Another person pauses to adjust their sleeves, belongings, or clothing, then continues. The other two can naturally pass each other, each going their own way. All their actions are unrelated to the two main characters. They must not look at the main characters. They must not stop because the main characters stop. They must not turn around because the main characters speak. They must not synchronize their movements for comedic effect. No visible background character can be completely frozen for an entire 10 seconds. Mid-range air layer: Fog, low clouds, or other reasonably existing air volume in the current reference world, continuously moving around real buildings and terrain. Fog must be able to be obscured by physical buildings. It should be briefly invisible after entering behind a building. It should then reappear from the other side according to the actual spatial relationships. It must never pass directly through physical structures. Foreground layer: During the actual movement of the camera, a foreground element consistent with the current reference world—fog, vegetation, banners, building edges, or other reasonable objects—shortly passes close to the lens. This makes the audience clearly feel that the camera is inside the space. If the reference environment contains water, damp stone, metal, jade, or other reflective surfaces, their reflections must continuously change with the camera position and viewing angle, and cannot be fixed as if drawn on a background image. All environmental movements share the same weather system and uniform wind direction. Character A, Sword Immortal Senior Sister: 25-30 years old, East Asian woman, oval face, fair complexion, dark almond eyes, long black hair half-up and secured with a white jade hairpin, tall and slender, wearing a white embroidered silk Hanfu, semi-transparent layered wide sleeves, silver waist belt, jade pendant, and white cloth boots. Character B, Junior Sister: 20-25 years old, East Asian woman, round and lively face, black hair braided, petite figure, wearing a light green linen Hanfu, dark belt, wooden hairpin, and black cloth shoes. 0-5s Panoramic or Long Shot – Main Character Track: The camera begins a realistic forward and slightly lateral push-in track from the interior of the reconstructed 3D world. Foreground space or air layers briefly pass near the lens. Mid-ground buildings create a clear parallax shift relative to the extremely distant background. The two walk side-by-side normally. Suddenly, the same senior swordswoman says very seriously, "Today we'll practice concentration." The same junior sister turns slightly to look at her. The senior swordswoman continues, "Whoever laughs first loses." The junior sister immediately puts on a completely expressionless face. Both stop simultaneously and turn to face each other. This entire scene only occurs because the senior swordswoman proactively suggests practicing concentration. It must not be triggered by any background changes. 0-5s Environmental lifecycle: While the two are speaking, a huge cloud in the distance continues to slowly migrate in its original direction. A distant character continues to cross a distant platform laterally. Another character continues to move along a real road or steps. Mid-range fog continues to flow at its original speed. Locally stable wind fields continue to produce small movements of clothing, vegetation, or hanging objects. The background cannot suddenly move when the dialogue begins. The background cannot stop when the two stop. 5-10s Mid-range two-person shot/cowboy scene—protagonist storycycle: The two face each other, about an arm and a half apart. Both try to maintain completely expressionless faces. No magic. No sword drawn. The environment is irrelevant. To break her senior's composure, the junior sister makes only a tiny movement: one cheek puffs out very slightly for about half a second, then immediately returns to normal. The sword-wielding senior almost reacts, but forces herself to remain calm. She raises one eyebrow with a tiny movement. The junior sister's lips almost curl upwards, but she immediately lowers them. The sword-wielding senior's lips also begin to tremble almost imperceptibly. Neither of them can truly laugh during this segment. The camera moves slowly and realistically in a small arc of about 15–20 degrees around the two characters. The camera must actually change position. Digital zoom cannot be used to simulate surround movement. This creates realistic parallax at different speeds in the near, middle, and far background. The comedy comes solely from the micro-expressions and restraint of the two characters. For 5-10 seconds, the environmental life track and the background character track continue to run independently: after the characters enter the medium close-up, the background world must not be turned into a static scene just because the shot begins to emphasize the face. Multiple clear motion sources must still be preserved at different depths. A figure in the distance continues walking forward, naturally obscured by existing buildings as they move. Another background figure emerges naturally from another previously obscured space and continues their path. The volume of air continues to move behind and between buildings. As the camera moves in an arc, the mid-ground structures and distant spaces continue to create different parallaxes. Clouds, fog, reflections, and distant figures must not stop moving even when the protagonist is making micro-expressions. All background movements are out of sync with the protagonist's rhythm. Ending transition: 9.5-10 seconds, the two protagonists' bodies are kept as stable as possible to facilitate a smooth transition. However, only the characters must be stabilized, not the world frozen. The Sword Immortal Senior Sister and Junior Sister are still facing each other. Both are clearly about to burst out laughing but are still forcibly maintaining a serious demeanor. The characters' center of gravity, gaze, and expressions are clear. The camera gradually stabilizes. At this point, the distant clouds can still be seen continuously moving. At least one distant figure is in the process of walking, not standing as a static human-shaped sticker. The mid-ground air is still in motion. At least one reflection or lighting condition is still continuously evolving. This precise state of "stable characters, but the world still in motion" serves as the transitional state for the second segment. It features a 16:9 landscape aspect ratio, native synchronized Mandarin dialogue, restrained background music, and realistic footsteps, fabric sounds, distant footsteps, wind sounds, and ambient sounds. The screen includes two main female characters, while also allowing for smaller, completely independent figures in the distance. No subtitles are generated, and no modern elements are present. Negative (paragraph 1 independent): blurry, bad quality, low quality, low resolution, noisy, jpeg artifacts, watermark, text, subtitles, error; deformed, mutated, bad anatomy, poorly drawn hands, bad composition, out of frame, disfigured; inconsistent protagonist identity, changing clothes, face morphing, hairstyle change; static background plate, frozen scenery, frozen background people, background people standing motionless for the whole clip, living people rendered as landscape texture, duplicated background extras, background extras staring at protagonists, background extras synchronizing with protagonists, background extras reacting to dialogue, background extras reacting to comedy; flat 2D reference image, animated wallpaper, camera sliding over a photograph, fake digital zoom, fake parallax, foreground midground and background moving at identical speed, no occlusion, no object permanence, fixed cloud texture, static cloud sea, frozen mist, static reflections, painted reflections, static distant people, lifeless surrounding architecture, environment stopping during dialogue, environment stopping during close-up, environment freezing When the protagonists stop, all environmental movement begins simultaneously, synchronized with plot beats, dramatic wind caused by dialogue, cloud change caused by laughter, lighting change used as punchlines, environment creation or story solving, camera following weather instead of protagonists, random new landmarks, impossible geometry, background replacement, teleporting extras, modern elements, and glitching cuts. The second segment is generated as an independent extended video. If the current Seedance workflow allows video continuation or video reference, the complete first segment video itself is prioritized as the temporal motion reference, while the last frame of the first segment is used as the opening visual state of the second segment. It is absolutely unacceptable to simply regenerate a "similar" location. The same world must continue moving forward along the timeline established in the first segment. The same sword-wielding senior sister and junior sister must be fully inherited, including identical faces, hairstyles, body proportions, clothing, accurate positioning, the suppressed laughter expression at the end of the previous segment, gaze, camera height, camera axis, and lens perspective. The motion phase of the living world itself must also be inherited. The background characters who were walking at the end of the first segment continue walking from their previous position and direction. They cannot return to the starting point. The clouds continue to migrate, maintaining their previously established direction and speed. The flowing fog continues its movement from the actual spatial state at the end of the previous segment. Reflections continue to evolve. It cannot suddenly revert to the state at the beginning of the first segment. The permissions of the three tracks remain completely unchanged: Main character plot track = 100% scriptwriting rights. Environmental life track = continuous physical movement rights, 0% scriptwriting rights. Background character track = independent living rights, 0% main character interaction rights. The background cannot cause the two characters to laugh. The environment cannot decide their wins or losses. Background characters cannot help complete the punchline. 10-15s close-up of two characters/restrained arc movement—Main character plot track: directly from the exact expressions of the two characters trying to suppress laughter at the end of the first segment. Both continue to try not to laugh. The same junior sister changes her strategy. She suddenly adjusts her posture to be extremely dignified, and then very seriously imitates the aloof, calm, and serious expression of the sword immortal senior sister. The same senior swordswoman immediately recognized that she was being imitated. A very slight change occurred in her nostrils. She held it in. Seeing this subtle reaction, the junior sister's lips began to tremble more noticeably. The senior swordswoman then noticed the junior sister struggling to suppress her laughter. So now it became: both of them were trying not to react to the other's effort not to laugh. They used only very small micro-movements, gradually escalating: a raised eyebrow, a suppressed breath, a slight tremor of the lower lip, an almost imperceptible shoulder tremor. No exaggerated ugliness. No large comedic movements. The environment was completely uninvolved in the humor. At approximately 14.5 seconds, both of them finally broke down at the exact same moment, simultaneously letting out a short, genuine laugh. This laugh could only come from the mutual expression of their reactions. For 10-15 seconds, the environmental lifeline continued simultaneously: even when the camera moved into close-up of the characters, the space behind them must still have clearly visible movement. One background character who was already walking in the first segment continued their path and was then naturally obscured by the existing architecture. Another background character passed through a deeper layer of space along a different path. A third background character performs a very small, unrelated action, then continues walking. The distant clouds continue to shift. The fog continues to move. Reflections continue to change with subtle shifts in the camera angle. The moment the two laugh: there can't be a sudden gust of wind. There can't be a sudden switch on the lights. The background character can't turn their head. The fog can't suddenly accelerate. There can't be any synchronized reaction from the background. 15-20s Mid-to-long shot ending—Protagonist's storyline: After a brief laugh, the two immediately regain their seriousness. The junior sister asks, "A tie?" The sword-wielding senior sister thinks for a second: "Let's start again." The junior sister nods seriously. The two solemnly restore their extremely formal expressions. Then they simultaneously turn back to the front. They continue walking forward side-by-side normally. After taking two steps: The junior sister casually glances at her senior sister. Unexpectedly, the same sword-wielding senior sister is already glancing at her. Their eyes meet. A very subtle smile reappears on both their lips. But this time, neither speaks. Don't add a third punchline. Finally, maintain an observational lingering effect. 15-20s Camera Track and Environmental Lifeline: As the two characters resume walking, the camera transitions to a soft three-quarter profile follow shot. The camera movement is still solely due to the characters walking. However, the camera must realistically create spatial displacement. A foreground layer, consistent with the reference world, naturally sweeps across the frame from one side. The two main characters are in the midground. In the deeper space, at least two autonomous background characters remain active. One character passes through a misty area, partially obscured by the air layer, and continues moving. The other character passes through at a significantly different speed at different depths. The vast, distant landscape moves very slowly relative to the others. The clouds and fog completely inherit the original direction of movement from the first segment and do not restart. If reflective surfaces exist in the scene, the highlights slide continuously as the camera's lateral position changes. After the characters' final eye contact and laugh, the background movement continues. Last 0.5 seconds: The two main characters are still walking forward. Simultaneously, the audience can still clearly see multiple independent sources of movement from non-main characters. The laugh has ended, but the world has not. 16:9 landscape mode, native synchronized Mandarin dialogue with realistic short laughs, precise lip-syncing, seamless transition with the first segment of the video, realistic camera parallax, autonomously moving background characters, continuous air movement, physically plausible dynamic reflections and ambient sound. Two main female characters, while also allowing for small background characters. No subtitles generated, no modern elements included. Negative (independent paragraph 2): blurry, bad quality, low quality, low resolution, noisy, jpeg artifacts, watermark, text, subtitles, error; deformed, mutated, bad anatomy, poorly drawn hands, bad composition, out of frame, disfigured; inconsistent protagonist identity, changing clothes, face morphing, hairstyle change; background reset between clips, background motion restarting from zero, background extras returning to starting positions, frozen background extras, static distant people, living people rendered as scenery texture, duplicated extras, disappearing extras without occlusion, background extras watching protagonists, background extras reacting to laughter, background extras laughing with protagonists, synchronized background choreography; flat background plate, static reference image, animated wallpaper, fixed cloud sea, frozen cloud structure, frozen mist, static reflections, painted reflections, fake parallax, digital zoom instead of camera translation, foreground midground background moving at identical speed, environment freezing in close-up, environment freezing during dialogue, environment freezing when protagonists laugh, environment freezing after punchline, environmental event causing laughter, wind gust synchronized with laugh, cloud burst synchronized with joke, sunlight burst at punchline, mist revealing something at story beat, background solving the contest, camera following environmental motion instead of protagonists, new random architecture, scenery replacement, impossible geometry, extra foreground protagonists, modern elements, glitching cuts