prompt.txt

subject_definitions:
<Subject 1> is the young Chinese woman in <Picture 1> (front view): fair skin, dark brown hair in a high ponytail with loose strands, wearing a white short-sleeve school polo shirt with blue collar and blue side panels, and dark navy track pants with white side stripes. <Picture 1> is the front-view identity anchor for <Subject 1> - use it for her face, hairstyle and costume.
<Picture 2> is a character turnaround sheet showing the SAME single woman from three angles (front, side, back) - the three figures are ONE person, not three different people. Use it only for costume details on the sides and back.
<Subject 2> is the young woman in <Picture 3> (front view): twin dark ponytails, black cat-ear hair accessories, a white short-sleeve T-shirt with black trim and a small cartoon print, black shorts, and white over-the-knee stockings. <Picture 3> is the front-view identity anchor for <Subject 2>.
<Subject 3> is the environment in <Picture 4>: a vast open white-jade viewing terrace floating above a sea of clouds, an ancient pine tree reaching in from the left, a bronze incense burner on a carved balustrade at the right, distant palace roofs emerging from mist, cool blue-grey sky with warm golden sunset light from the right. <Picture 4> is the scene anchor.
<Picture 5> is a POSE, STAGING AND CAMERA REFERENCE ONLY. It defines the camera height, the shot size, where each person stands and the exact body poses to reproduce frame by frame. The figures visible inside <Picture 5> are featureless 3D mannequins: they must NOT appear anywhere in the output and must never be copied. Replace them completely with <Subject 1> and <Subject 2>, keeping <Subject 1>'s and <Subject 2>'s own faces, hair and costumes.
All human subjects are photorealistic real people - not 3D mannequins, not anime, not illustrated.

summary:
[reference generation] A 5-second single continuous take with no cut. Reproduce the camera framing, staging and body poses of <Picture 5>, but performed by <Subject 1> and <Subject 2> on the jade terrace of <Subject 3>. <Subject 1> stands centre facing the camera with the right arm raised and the elbow bent, hand at shoulder height, the left arm relaxed at her side; she holds this gesture and sways slightly with the beat. <Subject 2> stands to her left, bent forward at the waist with the head lowered and both arms hanging down. The camera is static at chest height. The cloud sea drifts behind the balustrade. Nothing enters or leaves the frame.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - face, high ponytail, white-and-blue school polo and navy track pants match <Picture 1>; her raised-arm pose follows <Picture 5>.
<Picture 1> ([Shot 1] identity): fully_preserved.
<Picture 2> ([Shot 1] costume reference): partially_preserved - costume details only; the three views are one character, never a group.
<Subject 2> (appears in [Shot 1]): fully_preserved - twin ponytails, cat-ear accessories, white T-shirt, black shorts, white over-knee stockings; her bent-forward pose follows <Picture 5>.
<Subject 3> (appears in [Shot 1]): fully_preserved - terrace, pine, balustrade, incense burner and cloud sea match <Picture 4>; only the clouds drift.
<Picture 5> (appears in [Shot 1]): pose_reference_only - camera framing, staging and body poses are reproduced; the mannequin figures themselves are discarded and are NOT rendered.
<Picture 5> mannequins and overlays: not_referenced - the red and white mannequins, and any text, logo or watermark, must never be reproduced.

detailed_description:
[Shot 1] Photorealistic cinematic footage. Horizontal 16:9. One unbroken take for the full 5 seconds.
00:00-00:05.000: static shot at chest height, matching the framing, shot size and staging of <Picture 5>. <Subject 1> stands at the centre of the jade terrace of <Subject 3>, framed from head to mid-thigh, in the same pose as the central figure of <Picture 5>: the torso faces the camera, the right arm is raised with the elbow bent and the hand near shoulder height, the left arm hangs relaxed. She keeps the raised hand and moves subtly with the beat - the shoulders sway, the raised hand rotates slightly, the head tilts a little, the ponytail swings. <Subject 2> stands to her left in the same bent-forward posture as the left figure of <Picture 5>: waist bent, head lowered, both arms hanging down; she holds the pose and sways gently. Only these two women are present.
The featureless 3D mannequins of <Picture 5> must never appear; only their poses, staging and camera framing are reproduced. No text, no subtitles, no watermark, no logo, no UI overlay anywhere in the frame. No 3D mannequin, no anime style, no illustrated look, no extra person, no duplicated person, no camera shake, no green screen.

overall_soundscape:
High-altitude wind, faint cloth fluttering, soft footsteps on stone. No dialogue.

non_diegetic_music:
A steady mid-tempo electronic dance beat with a soft kick on every beat and a light synth pulse, consistent from start to finish.
下载此文件