Generate a 15 second video, with yourself, sharing the last bite with a stray animal.
Generate a 15 second video, with yourself, sharing the last bite with a stray animal.
10-second 16:9 ultra-photoreal cinematic architectural commercial. Follow the same mechanical assemble-to-hero beats. @image1 is the completed building and site — photocopy that architecture, terraces, green roofs, hedges, street, cars, and facade. Same materials as the photo. Do not invent a different development. Only gloved hands in matte-black gloves. Bright hard sunlight, clear deep-blue sky. 0.0–0.6s: extreme macro, a finger draws a straight line through sand, revealing a concrete foundation. 0.6–1.3s: a hand places one concrete core; columns and structural grid assemble outward. 1.3–2.0s: two fingers pull upward; floors, slabs and columns rise and lock. 2.0–2.7s: a hand slides the upper volume sideways, forming terraces, setbacks and deep balconies; short orbit. 2.7–3.4s: a hand sweeps the frame; stone, metal, panoramic glass and balcony details install. 3.4–4.1s: fingers click the final panel; macro glide across stone + metal + glass. 4.1–4.8s: a finger draws a path through sand; paving assembles into paths, plaza and driveway. 4.8–5.5s: a hand plants one mature tree; surrounding trees, lawn and shrubs appear to match @image1 landscaping. 5.5–6.2s: fast hand sequence places bench → architectural light → a dark luxury SUV like the one in @image1. 6.2–7.0s: a finger draws an arc; ground lowers into a shallow reflecting pool and fills. 7.0–8.2s: fingers adjust the final tree; camera pushes into leaves; scale shift from miniature to real-world; hands disappear. 8.2–10.0s: camera emerges under the real tree, flies low over paving, rises along the facade, orbits the corner, ends on a wide HERO SHOT matching @image1. Physical mechanical transformations only: sliding, unfolding, assembling, locking, realistic mass and shadows. No magic, glow, holograms, toys, cranes, workers. No extra on-screen captions. AUDIO: quiet construction clicks, wind, no voiceover.
A winter coat closet, rendered as a stylized animated feature. A tiny steam locomotive bursts out of a hanging coat's pocket in a spray of snow, a mouse conductor leans from the cab and calls the last train, and pocket seventeen opens onto a snowy tunnel glowing blue. Follow the train through and you land on the platform of a miniature railway station hidden inside the coat's quilted lining, where a mouse child is waiting with one lost mitten and no match for it. Pick or create your avatar, build a turnaround sheet so the character holds through every shot, lay the twelve beats out as a storyboard, then render the finished film.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 10 seconds. AUDIO: Generate synchronized natural audio — ambient light rain on a coat, a rising whoosh as time freezes, near-silence during the freeze, then a sharp splash and a delighted laugh. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. Photocopy this person. High bun with gold-and-green butterfly clips, teal velvet embroidered jacket, geometric earrings. Hair stays in that bun for every frame. PHOTOREAL LIVE-ACTION: Real camera, real tiny human, real jacket pocket. Not CGI, not a figurine, not a doll, not a sticker, not a composite. SUBJECT LOCK: Exact face, features, skin, build and gender from @image1. Only scale changes. Do not beautify or restyle. FIRST-FRAME READ — HIGHEST PRIORITY AFTER IDENTITY: A first-time viewer must understand in under one second: a miniature person is inside a clothing pocket. Camera sits OUTSIDE the garment, chest-to-hip height, looking down and in through the pocket mouth. Never from inside the well. Never looking straight up at the sky. Never a zipper tunnel. Never a fabric cave. Never a bag. In EVERY frame, including frame 1 and the freeze-orbit, keep all three in view at once: 1) the COAT — rain-wet outer fabric, seam, buttons or zipper of the garment running beside the pocket, enough torso that it is obviously a jacket worn by a different full-sized adult (no face on the wearer, never a second copy of @image1) 2) the POCKET as a 3D pouch sewn onto that coat — welt or flap, folded hem, bartack stitching, the U-shaped opening 3) the TINY SUBJECT down inside that pouch — a few centimetres tall, occupying at most one-fifth of the frame height so scale is obvious. She stands on the inner fabric floor; walls surround her; she is not perched on the pocket like a brooch. If the pocket silhouette would leave frame, pull the camera back. Frame wide. SHE IS INSIDE THE POCKET BAG — NOT ON IT: legs and hips on the inner floor, cloth folds wrap her, contact shadow, she occludes the weave behind her. Outdoors, light drizzle. Pale daylight and rain enter FROM ABOVE through the opening. The sky is a strip above the pocket mouth, not the whole frame. Same light on her skin, the droplet and the fabric. cameraFixed unset. CONTINUOUS MOTION except the designed freeze. 0-1s HOOK — already mid-action, WIDE: rain-wet jacket pocket, tiny woman already inside the pouch shielding her eyes, off-balance. Coat, pocket hem, stitching, and her full tiny body all readable. 1-3s: One giant raindrop falls from the sky strip INTO the pocket opening toward her, growing larger. Camera still outside, pocket still readable. WOW — TIME FREEZE + ORBIT: Drop hangs inches from her, spherical, refracting light. She frozen mid-flinch. Slow 180-degree orbit around droplet + tiny person. During the orbit the coat, pocket opening and hem MUST stay in frame — do not push into a fabric cave. Freeze ~2s. TIME RESUMES — splash past her, spray on inner fabric. 6-8s: Half-drenched, she laughs, wipes her face, sees a puddle on the fabric floor with her reflection. Still down inside; coat and pocket hem still in frame. 8-10s: Scoops puddle in cupped hands, looks into camera, grins, still dripping, still inside the cavity, jacket pocket still readable around her. SCALE: stitching, hem and raindrop are huge next to her; she is tiny next to the coat. No on-screen text, no extra watermark, no subtitles, no gym, no vanity lights.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 8 seconds. AUDIO: Generate synchronized natural audio for this scene — a low buzzing hum that builds as the vibration intensifies, the coin and folded receipt rattling against the fabric floor, a startled yelp, and a relieved laugh at the end. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Exact face, facial features, hairstyle, skin tone, build and gender presentation stay unchanged — only their scale changes. Do not beautify, restyle or alter identity. Photocopy @image1. Face, hair, skin, the high blonde ponytail with the long braid over one shoulder, the green satin button-down, and the gold necklace from the still stay. POCKET READ — MUST BE OBVIOUS FROM FRAME 1: This is a real jeans FRONT POCKET on blue denim jeans worn at the hip of a different, full-sized adult (never the subject’s face). A first-time viewer must instantly recognise it as a clothing pocket — not a fabric cave, not a bag, not a tunnel of cloth. Camera sits OUTSIDE the garment, slightly above hip height, looking down and in through the pocket opening. In every frame keep visible at once: - the jeans: waistband, a belt loop, bartack stitching, rivet, denim weave - the pocket as a 3D pouch sewn onto the jeans, with a clear U-shaped opening and a folded pocket-edge hem - the tiny subject inside that pouch, a few centimetres tall, with the pocket bag huge around them - a coin and a folded paper receipt resting on the fabric floor nearby, enormous next to them Do not crop so tight that only fabric walls remain. If the pocket silhouette would leave the frame, pull the camera back. CONTINUOUS MOTION REQUIRED: Real, visible physical motion for almost the entire 8 seconds — a clear sequence of distinct actions, never one static pose held for seconds at a time. SCENE: The subject, shrunk to a few centimeters tall, is inside that jeans pocket. A coin and a folded receipt rest on the fabric floor nearby. Indoors, soft natural window light from one side so fabric texture and the subject both read clearly. The wearer of the jeans is a different full-sized adult — never give them the subject’s face. No gym, no banners, no second copy of the subject. 0-1.5s: Wide enough to read the jeans pocket on the hip. Tiny subject sits calmly inside the pocket bag, idly straightening the folded receipt beside them, at ease in their little space. Waistband and belt loop still in frame. 1.5-3s: A phone begins ringing in a nearby pocket — the entire fabric floor starts buzzing and shaking rhythmically beneath them. Startled, the subject grabs onto a fold of fabric to keep their balance as the coin and receipt rattle and skitter across the floor. The U-shaped pocket opening stays readable. 3-5s: The vibration intensifies. The subject stumbles and slides across the shaking fabric, arms wheeling, trying to stay upright as the whole pocket pouch trembles around them — a genuinely comedic struggle against the motion. Jeans waistband still in frame. 5-6.5s: The vibration cuts off abruptly. The subject, thrown off-rhythm, stumbles forward and catches themselves against the fabric wall, dizzy for a beat, blinking. Pocket silhouette still readable. 6.5-8s: They shake off the daze, laugh, straighten up, and peek upward through the pocket opening at the now-quiet world above, giving a small relieved thumbs-up toward camera. Waistband still readable. SCALE REALISM: The pocket’s real stitching, fabric weave and denim texture must read at true size next to the tiny subject — the coin, the receipt and every fold of fabric should look enormous relative to them. Match lighting and color grade between the subject and the fabric throughout so the whole shot reads as one real photograph, not a composite. Keep the camera framing wide enough that the shaking, the stumble and the recovery all stay clearly visible in frame. No on-screen text, no watermark, no subtitles.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 8 seconds. AUDIO: Generate synchronized natural audio for this scene — soft fabric rustling as the subject climbs and grips the pocket edge, a sharp metallic clink when the coin drops in, a startled gasp, and a genuine laugh at the end. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Exact face, facial features, hairstyle, skin tone, build and gender presentation stay unchanged — only their scale changes. Do not beautify, restyle or alter identity. Photocopy @image1. Face, hair, skin, and the white collar/shirt from the still stay. POCKET READ — MUST BE OBVIOUS FROM FRAME 1: This is a real jeans FRONT POCKET on a pair of blue denim jeans worn at the hip of a different, full-sized adult (never the subject’s face). A first-time viewer must instantly recognise it as a clothing pocket — not a fabric cave, not a bag, not a tunnel of cloth. Camera sits OUTSIDE the garment, slightly above hip height, looking down and in through the pocket opening. In every frame keep visible at once: - the jeans: waistband, a belt loop, bartack stitching, rivet, denim weave - the pocket as a 3D pouch sewn onto the jeans, with a clear U-shaped opening and a folded pocket-edge hem - the tiny subject inside that pouch, a few centimetres tall, with the pocket bag huge around them Do not crop so tight that only fabric walls remain. If the pocket silhouette would leave the frame, pull the camera back. CONTINUOUS MOTION REQUIRED: Real, visible physical motion for almost the entire 8 seconds — distinct actions, never one static pose held for seconds. Indoors, soft natural window light from one side. The wearer of the jeans is a different full-sized adult — never give them the subject’s face. 0-1.5s: Wide enough to read the jeans pocket on the hip. Tiny subject sits inside the pocket bag, looking up at the opening and the light. Waistband and belt loop still in frame. 1.5-3.5s: They grip the folded pocket-edge hem with both hands and climb, hoisting themselves up until perched on the rim of the pocket opening, legs dangling over the outside of the jeans, catching their breath. The U-shaped pocket opening stays readable. 3.5-5s: The person wearing the jeans shifts their weight; the whole pocket pouch sways. The subject grips the hem tightly, body rocking, still clearly inside a jeans pocket on a hip. 5-6.5s: A coin drops in from above through the pocket opening and lands with a clatter beside them in the pouch — the subject flinches back, startled, both arms raised. The coin is enormous next to them. Pocket opening still visible. 6.5-8s: They relax, walk over and pick up the coin with both hands, straining against its size, then hold it up over their head in triumph, laughing — still standing inside the jeans pocket, waistband in frame. SCALE REALISM: Stitching, denim weave, belt loop and coin are true size next to the tiny subject. Match lighting so it reads as one photograph, not a composite. No on-screen text, no watermark, no subtitles.
SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Photocopy their exact face, facial features, hairstyle, skin tone, build and gender presentation. Do not beautify, restyle, de-age or alter identity. Only their scale changes. @image1 is the subject. Keep the navy square-neck top, gold coin pendant necklace, gold signet ring and gold chain bracelet exactly as shown. Only size changes: they are a few centimetres tall. CONTINUOUS MOTION REQUIRED: Real, visible physical motion for almost the entire 8 seconds — not a static pose. Something is always moving: the giant finger descending, the subject dodging, the grab, the swing, the landing. Do not default to a calm, mostly-still shot. SCENE: The subject, shrunk to a few centimetres tall, stands on an open human palm — a different, full-sized adult hand, not the subject’s own. The hand is held at roughly waist height, palm up, fingers relaxed and slightly spread, indoors with soft natural window light from one side so the skin texture and the subject both read clearly. Photoreal. One continuous real photograph, not a composite. SCALE REALISM: The palm’s skin texture, lines, and the finger’s fingerprint ridges must read at true full size relative to the tiny subject — every texture detail on the hand looks enormous next to them. Match lighting and colour grade between the subject and the hand throughout. Keep the camera framing wide enough that the entire arc of motion (finger descending, the grab, the swing, the landing) stays visible in frame. Never open on the attached studio still or a grey headshot. No on-screen text, no watermark, no subtitles. No other people except the giant hand that holds them. 9:16. TOTAL DURATION: 8 SECONDS. 0-1s: Subject stands calmly in the centre of the palm, weight settled, looking around at their surroundings. Full body of the tiny subject is readable. Palm lines look enormous next to them. 1-3s: A giant index finger slowly descends from directly above the frame toward the subject — clearly visible entering from the top, moving at a deliberate, readable speed, not too fast to track. As it nears palm height, the subject notices, reacts with a startled flinch, and dodges sideways just in time, the fingertip passing close beside them. 3-6s: Before the finger retreats, the subject changes their mind mid-motion — they lunge forward and grab onto the fingertip with both hands, fingers wrapped around the tip. The giant hand lifts slightly, carrying the subject a short distance up and sideways through the air in a small arc — a genuine swinging motion, not just a lift. The subject’s legs kick slightly, hair and clothing trailing with the motion, face showing visible thrill and a wide grin. 6-8s: The finger lowers the subject back down onto the palm. They let go, stumble half a step to find their footing, then straighten up and give the giant finger a playful pat with one hand, laughing. Original instrumental music throughout — cinematic, no lyrics, no known songs.
Create a 15-second ultra-realistic cinematic 16:9 video. Sound ON. If Montenegro were a chocolate bar. ATTACHED IMAGES ARE THE COUNTRY — photocopy them, do not invent a generic coast or Alps. @image1 is the REAL Sveti Stefan: tiny Adriatic island packed with stone houses and terracotta roofs, connected by a narrow causeway/isthmus, turquoise water, pink-orange stone. @image2 is the REAL Durmitor: high pale-grey limestone peaks, jagged, alpine, sparse snow, not rolling green hills. OPENING: dark chocolate bar, red-gold wrapper, dark table. Hand snaps it. The break REVEALS these two places, unmistakably: first Sveti Stefan as in @image1 (island + isthmus + packed roofs), then Durmitor as in @image2 (high karst massif). Melted chocolate = bay. Crumbs = karst. Finish aerial: Sveti Stefan in water in front, Durmitor peaks behind. Broken chocolate at edges. No Japan, no Fuji, no generic Cinque Terre, no on-screen text. AUDIO: snap, sea, mountain wind. No vocals. Duration 15s. Aspect 16:9.
Create a first-person POV viral video showing the magic of architectural design coming to life. Real hands in black leather gloves hold up an architectural blueprint of a sleek modern villa, then pull it away to reveal the villa magically assembling in real-time. Concrete walls, steel frames, and glass panels rapidly materialize with cinematic effects, dust clouds billowing as construction happens at impossible speeds. The blueprint elements transform into real structural components. Set on a bright sunny construction site with clear blue skies.
CHARACTER LOCK — HIGHEST PRIORITY: The person in every shot is exactly the person in @image1. Preserve their facial identity, facial structure, eye shape and spacing, nose, mouth, exact skin tone, hairline, hairstyle, hair length, texture and colour, freckles and distinctive marks, apparent age, build and body proportions — identically, shot to shot. Preserve their gender presentation exactly as shown; do not masculinize, feminize or reinterpret it. Do not beautify, slim, de-age or restyle them. The attached image outranks every word below it: wherever the image and the text disagree, follow the image. Frame 1 is already in the scene — never a studio portrait, grey backdrop, or the attached still as an opening plate. Use they/them throughout. No other people. Create a 30-second 16:9 2D hand-painted storybook animation — not photoreal, not live-action, not 3D CGI. Visible brush, soft painted edges, night-forest colour. The person in the attached reference is the only person. They are a fairy in this scene: same face as @image1, plus delicate translucent wings. OPENING — MANDATORY: Frame 1 is already a close-up of a hanging bell-shaped glowing flower on a mossy log, bare feet stepping into frame. Handheld-feeling but painted. Do not open on a studio portrait, grey backdrop, still headshot, or the attached image as a plate. Wardrobe, locked, gender-neutral: hand-knitted blue tunic, loose dark trousers, plain blue wool cap, small brown leather pouch on a strap, barefoot, translucent fairy wings. Fit these clothes and the wings to the attached person's body. No dress, no skirt, no gown. Hair stays as in the attached reference — mist can dampen it, do not recut or recolour it. Setting: enchanted night forest. Deep blue trees, mossy log, many hanging bell-shaped flowers on thin vines, floating pink purple and blue glow-lights, fog, tiny sparkles. One person only. 0-6s — hook: Bare feet on wet moss. A large pale-lilac bell flower hangs in the foreground, glowing softly. They step along the log toward it. Wings flutter. 6-14s — approach: They walk the log in the knitted tunic, trousers and cap, pouch at their side. Painted bokeh lights behind. Slow move toward the hanging flower. 14-22s — the key: They crouch, open the pouch, take out a small ornate golden key. Look at the flower. Reach toward it. 22-30s — glow: They touch the flower with the key. The bell lights from inside, sways, brighter. Hold on the glowing flower as they watch. No on-screen text. Cut. Diegetic forest night only: soft wing flutter, faint chime when the flower lights, breeze. No music, no speech, no subtitles, no logos, no watermark. Duration: exactly 30 seconds. Aspect ratio: 16:9.
Ultra-photorealistic commercial. Live-action macro tabletop cinematography throughout: real optics, true shallow depth of field with a fast macro lens, natural motion blur, fine dust in the air, warm practical light spilling from the model's windows against cool ambient fill. Real materials at every scale — timber grain, laid shingles, weathered stone, cured resin, glass, and the fine-grain matte weave of the gloves. No stylisation of any kind. Gloved hands in matte-black technical gloves build a handcrafted diorama on a dark graphite workbench: a steep rocky outcrop carrying a tall alpine mill house in weathered timber and stone, steeply pitched shingled gable roof, a timber waterwheel on its flank, a modelled stream bed running down under a plank footbridge to a pool, and a vertical glass cutaway slicing through that pool. Hold the model identical in every shot — same architecture, proportions, materials and colours, no drift and no reinterpretation. THE MODEL IS A STATIC, INERT OBJECT ON A BENCH. It is never animated and never comes to life. Through every miniature shot the only things that move are the gloved hands, the resin while it is being poured, and the camera. The model's water is cured resin and is rock-solid and mirror-still; the sculpted whitewater is hard and frozen; the fish are cast pieces fixed in the resin; the waterwheel is seated in its bracket and stays motionless at all times. Hands assemble the model piece by piece on the workbench, every part installed physically: INSERT → ALIGN → SEAT → SET. Gloved fingertips set the first cut stone into the rock base with fine steel tweezers, CLICK. Then both hands lower the timber mill-house shell down onto the rock outcrop and seat it. Then one hand seats the timber waterwheel into its bracket while the other steadies it, and the hands withdraw leaving it still. Then a fingertip presses the last row of roof shingles down flat. Then tweezers plant a slender conifer into a crevice in the rock. Then a hand feeds the LED wire in under the base and the windows come on, warm amber from inside. Then both hands press the glass cutaway wall into its channel and run a bead along the seal, the dry pool floor behind it already laid with pale sand, boulders and small painted fish. Then a hand tilts a graduated mixing cup and a clear ribbon of resin runs down into the basin, spreading across the sand and levelling out to a glass-hard, perfectly flat surface, and a cloth polishes the outside of the glass — the camera pushes in on the motionless wheel until the frame is filled by it → MATCH CUT. Then that same wheel is the wheel of a real full-size mill above a river, now genuinely turning under falling water, river running white over the rocks, mist lifting, the camera tracking low and wide as the light goes warm and long across the valley. Give each beat the same weight and the same screen time — no beat rushed, none held longer than the rest. Extreme macro, overhead, POV and low tabletop angles. Fast hard cuts throughout the build, one continuous move only on the push into the wheel and the shot it lands in. One camera move per shot. Audio: fine mechanical CLICKs as parts seat, tweezers ticking on stone, a soft bead of sealant, resin running and settling, cloth on glass, then a hard cut into full-size river sound and the creak of a working wheel. The miniature must read as a real handbuilt static model of real architecture, never as a toy. No moving water in the model, no flowing stream, no rippling or sloshing resin once poured, no swimming fish, no turning wheel before the match cut, no drifting weed, no smoke, no self-assembly, no part moving on its own, no magic, no plastic toy surfaces, no cartoon or stylised rendering, no tilt-shift look on the full-size shot, no text, no captions, no subtitles, no logo, no watermark, no UI.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 10 seconds. AUDIO: Generate synchronized natural audio — rain ticking on a coat, fabric creak, a rising whoosh as time freezes, near-silence in the freeze, then a sharp splash and a delighted laugh. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. Photocopy. High bun, gold-and-green butterfly clips, teal embroidered velvet jacket, geometric earrings. Bun stays. Only scale changes. THEME — POCKET WORLD: The film is a miniature world INSIDE a real clothing pocket. A first-time viewer must get that in under one second. The pocket is both a place she lives and a pocket on a coat. HOW THE POCKET READS (mandatory, every frame including freeze-orbit): Camera OUTSIDE, chest-to-hip, looking down through the pocket mouth into the pouch. Never from inside a well. Never a zipper tunnel. Never looking straight up at sky. Never the gap between coat buttons. Never a bag, cave, or hallway of cloth. In frame at once, always: - a rain-wet JACKET on a different full-sized adult (torso only, no wearer face, never a second @image1) - a PATCH or WELT POCKET sewn onto that jacket — U-shaped opening, folded hem, bartack, a metal rivet the size of a boulder next to her - the TINY SUBJECT a few centimetres tall, at most one-fifth of frame height, standing ON THE INNER FABRIC FLOOR of that pouch - POCKET-WORLD PROPS on that same floor, huge beside her: a coin, a crumpled receipt, lint. These prove the floor is the inside of a pocket. She is IN the cavity, not on the pocket like a brooch. Legs on the inner floor, walls around her, contact shadow. The opening is a sky-window above her — grey drizzle and daylight come IN through it. Sky is a strip at the pocket mouth, not the whole frame. If the pocket silhouette would leave, pull back. PHOTOREAL. Match grain and light on her, the coin, the droplet and the fabric. cameraFixed unset. 0-1s HOOK — WIDE, already mid-action: rain-wet jacket, obvious patch pocket, tiny woman already down inside next to the giant coin, shielding her eyes. Instant read: pocket world. 1-3s: One raindrop falls from the sky-window through the opening toward her, growing. Coin and receipt stay on the floor. Pocket hem and rivet still in frame. WOW — TIME FREEZE + ORBIT: Drop hangs inches from her, spherical. She frozen mid-flinch. Slow 180 orbit. Coat, pocket hem, rivet and coin MUST stay visible — do not push into a fabric cave. ~2s. TIME RESUMES — splash, spray on coin and inner fabric. 6-8s: She laughs, wipes her face, sees a puddle on the fabric floor beside the coin, her reflection in it. Still inside the pouch; hem and rivet in frame. 8-10s: Scoops puddle, grins at camera, still dripping, still in the pocket world. No on-screen text, no extra watermark, no subtitles, no gym, no vanity lights, no button-placket standing shot.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 8 seconds. AUDIO: Generate synchronized natural audio — a low phone-vibration buzz that builds, coin and folded receipt rattling on denim, a startled yelp, a relieved laugh. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. Photocopy this person. PHOTOREAL LIVE-ACTION — HIGHEST PRIORITY AFTER IDENTITY: This must look like a real camera filmed a real tiny human inside a real jeans pocket. Not CGI, not a 3D figurine, not a doll, not a sticker, not a composite. Match grain, colour, and lighting between the person and the denim so they are one photograph. SUBJECT LOCK: Exact face, facial features, hairstyle, skin tone, build and gender from @image1. High blonde ponytail with the long braid over one shoulder, green satin button-down, gold necklace. Do not beautify or restyle. Only scale changes. SHE IS INSIDE THE POCKET BAG — NOT ON IT: The subject is a few centimetres tall and has climbed down INTO the pocket cavity. Fabric walls surround her. She sits on the inner fabric floor, not perched on the outside of the pocket like a brooch. Her legs and hips sink into the denim; cloth folds wrap her; she casts a contact shadow; she occludes the weave behind her. A real coin and a crumpled receipt share that same inner floor and are huge next to her. POCKET READ FROM FRAME 1: Camera slightly above hip height, looking down THROUGH the U-shaped pocket opening into the pouch. Keep in frame: blue-jeans waistband, one belt loop, bartack, copper rivet, denim weave, the folded pocket-edge hem, and the tiny person down inside. Instantly readable as a jeans front pocket on someone else’s hip — never her face on the wearer. Not a fabric cave, not a bag, not a tunnel, not a person glued onto the pocket front. LIGHTING: One soft indoor window from the side. The same light hits her skin and the denim. No separate beauty key on her. Pores, flyaways, fabric pills. Slight handheld micro-movement. CONTINUOUS MOTION: Real physical motion almost the entire 8 seconds. 0-1.5s: She sits inside the pouch, weight on the fabric floor, idly straightening the huge receipt. Denim around her. Waistband readable. 1.5-3s: A phone in another pocket starts buzzing. The whole pouch floor quivers. She grabs a fold. Coin and receipt hop and skitter. Denim ripples. 3-5s: Buzz intensifies. She stumbles and slides across the inner floor, arms wheeling, fabric throwing her — messy, physical, not posed. 5-6.5s: Buzz cuts. Momentum carries her into the inner wall. She catches herself, dizzy, blinking. 6.5-8s: She laughs, steadies, looks up through the pocket opening, small thumbs-up toward camera — still down inside the pouch, waistband in frame. SCALE: Stitching, rivet, coin and receipt are true size next to her. No gym, no banners, no second copy of her face, no on-screen text, no extra watermark, no subtitles.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 8 seconds. AUDIO: Generate synchronized natural audio for this scene — soft fabric rustling as the subject climbs and grips the pocket edge, a sharp metallic clink when the coin drops in, a startled gasp, and a genuine laugh at the end. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Exact face, facial features, hairstyle, skin tone, build and gender presentation stay unchanged — only their scale changes. Do not beautify, restyle or alter identity. Photocopy @image1. Face, hair, skin, and the white collar/shirt from the still stay. POCKET READ — MUST BE OBVIOUS FROM FRAME 1: This is a real jeans FRONT POCKET on a pair of blue denim jeans worn at the hip of a different, full-sized adult (never the subject’s face). A first-time viewer must instantly recognise it as a clothing pocket — not a fabric cave, not a bag, not a tunnel of cloth. Camera sits OUTSIDE the garment, slightly above hip height, looking down and in through the pocket opening. In every frame keep visible at once: - the jeans: waistband, a belt loop, bartack stitching, rivet, denim weave - the pocket as a 3D pouch sewn onto the jeans, with a clear U-shaped opening and a folded pocket-edge hem - the tiny subject inside that pouch, a few centimetres tall, with the pocket bag huge around them Do not crop so tight that only fabric walls remain. If the pocket silhouette would leave the frame, pull the camera back. CONTINUOUS MOTION REQUIRED: Real, visible physical motion for almost the entire 8 seconds — distinct actions, never one static pose held for seconds. Indoors, soft natural window light from one side. The wearer of the jeans is a different full-sized adult — never give them the subject’s face. 0-1.5s: Wide enough to read the jeans pocket on the hip. Tiny subject sits inside the pocket bag, looking up at the opening and the light. Waistband and belt loop still in frame. 1.5-3.5s: They grip the folded pocket-edge hem with both hands and climb, hoisting themselves up until perched on the rim of the pocket opening, legs dangling over the outside of the jeans, catching their breath. The U-shaped pocket opening stays readable. 3.5-5s: The person wearing the jeans shifts their weight; the whole pocket pouch sways. The subject grips the hem tightly, body rocking, still clearly inside a jeans pocket on a hip. 5-6.5s: A coin drops in from above through the pocket opening and lands with a clatter beside them in the pouch — the subject flinches back, startled, both arms raised. The coin is enormous next to them. Pocket opening still visible. 6.5-8s: They relax, walk over and pick up the coin with both hands, straining against its size, then hold it up over their head in triumph, laughing — still standing inside the jeans pocket, waistband in frame. SCALE REALISM: Stitching, denim weave, belt loop and coin are true size next to the tiny subject. Match lighting so it reads as one photograph, not a composite. No on-screen text, no watermark, no subtitles.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 8 seconds. AUDIO: Generate synchronized natural audio for this scene — soft ambient room tone, a light tap or brush sound on each landing on skin, the subject’s own breathing, small gasps and a genuine laugh at the end. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. If it is a person: their exact face, facial features, hairstyle, skin tone, build and gender presentation stay unchanged — only their scale changes. If it is a product/object: its exact shape, color, materials, proportions and any label or logo text stay exactly as shown, not regenerated — only its scale changes. Do not beautify, restyle or alter identity or branding in any way. CONTINUOUS MOTION REQUIRED: This shot must show real, visible physical motion for almost its entire 8 seconds — a fast sequence of distinct actions, never one static pose held for seconds at a time. SCENE: The subject, shrunk to a few centimeters tall, stands near the wrist of an open human palm — a different, full-sized adult hand, not the subject’s own — held at roughly waist height, indoors with soft natural window light from one side. 0-1s: Subject crouches at the base of the palm, looking up at the four fingers laid out ahead like stepping stones, visibly bracing to move. 1-2.5s: The nearest finger slowly curls down closer to palm height, and the subject leaps up onto its fingertip in one quick, athletic jump, arms out for balance. 2.5-4s: Without pausing, they leap again to the next fingertip — a faster, more confident jump this time, hair or clothing flicking with the motion. 4-5.5s: A third leap to the next fingertip — they land slightly off-balance, wobble, arms windmilling, and catch themselves just in time, a flash of surprise on their face. 5.5-8s: They plant their feet and launch into one final big leap off the last finger, arcing through the air. The giant hand moves quickly underneath and cups shut around them, catching them safely. The subject throws both arms up inside the curled hand in triumph, laughing. SCALE REALISM: The palm’s skin texture, lines and fingerprint ridges must read at true full size relative to the tiny subject — every texture detail on the hand should look enormous next to them. Match lighting and color grade between the subject and the hand throughout so the whole shot reads as one real photograph, not a composite. Keep the camera framing wide enough that all three jumps and the final catch stay clearly visible in frame. No on-screen text, no watermark, no subtitles.
SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Their exact face, facial features, hairstyle, skin tone, build and gender presentation stay unchanged — only their scale changes. Do not beautify, restyle or alter identity in any way. Use they / their / them only. @image1 is the subject. Keep their green satin shirt, high-waisted black trousers, long blonde high-ponytail braid, gold necklace, tan loafers and tan handbag exactly as shown. Only their size changes: they are a few centimetres tall. SCENE: The subject, shrunk to a few centimetres tall, stands on an open human palm — a different, full-sized adult hand, not the subject’s own. Photoreal close-up of that palm: natural skin, enormous creases and fingerprints next to the tiny loafers, fingers that dwarf the subject. Indoor warm light matching the green silk. Shallow depth of field. One continuous real photograph, not a composite. No gym, no brick wall, no banners, no other people in frame except the palm that holds them. SCALE REALISM: Palm skin texture, lines and fingers read at true full size relative to the tiny subject. Match lighting and colour grade between the subject and the palm. Frame close enough that the scale contrast is unmistakably visible — this is the single most important requirement. Never open on the attached still, a studio portrait, or a gym interior. No on-screen text, no watermark, no subtitles. 9:16. TOTAL DURATION: 8 SECONDS. Original instrumental music throughout — cinematic, no lyrics, no known songs. 0.0–2.0s MACRO PALM. Extreme close-up already on the open adult palm. The tiny subject stands in the centre of the palm, full body readable, a few centimetres tall. Palm lines look enormous next to their loafers. Fingers dwarf them. 2.0–5.0s SCALE HOLD. Slow push-in. The subject shifts weight slightly; the braid and tan bag read at miniature scale. Fingerprints and creases are landscape around them. Lighting matched. 5.0–8.0s CARE. The fingers curl a little inward as if cupping them safely. The tiny subject looks up toward camera. Hold the scale contrast. Music continues. End on the palm, not a cutaway.
【Generation Goal】 Generate a premium cinematic short film in which someone only a few inches tall crosses a worktable, finds a tear in an enormous jacket, threads a needle taller than they are, repairs the seam by hand, and is finally revealed riding in the breast pocket of the finished jacket as a giant wears it. 【Visual Style】 High-end cinematic 3D animation of polished feature-film quality, with realistic materials, detailed fabric, natural textures, physically based lighting, realistic shadows and subtle depth of field. Cloth reads as real wool and cotton with legible weave, fibre and thread detail and correct weight in every fold. Cloth, wood, metal and paper each read as their own material. Skin is soft and matte with subsurface warmth. Long soft shadows and rich material separation, with the lighting taken from @image2. Do not render a plastic, vinyl, toy or figurine look, cartoon physics, clay or stop-motion surfaces, photoreal live-action photography, ink outlines or line-art contours, cel shading, flat banded shadows, comic, manga, anime or webtoon styling, or vector or flat-illustration fills. 【Reference Asset Roles】 @image1 defines the character's face, hair, body proportions, clothing and accessories; take the identity and costume entirely from it and do not adopt its neutral background. @image2 defines the location and every object in it: the workshop, the worktable, the jacket, the needle, the spool of thread, the measuring tape, the tools and the lighting. It also defines scale — the size relationship between the character and every object is fixed by @image2 and never changes. Do not take any person from it. @image3 is the storyboard and governs shot order, framing, staging, character position and camera geography; read it left to right and then top to bottom, and do not adopt its panel grid, gutters or flat layout into the video. 【Subjects and Relationships】 There is one continuous character throughout, defined by @image1, only a few inches tall. Use they / their / them. They are never swapped, duplicated, replaced or altered in face, hair, clothing, proportions, colours or accessories in any shot. They work with calm, practised care and quiet pride. A second, giant person appears late — first as a hand and wrist entering frame, then as a body wearing the jacket. This giant is always framed so that no face and no head is visible: the hand shot shows only hand and wrist, and the final shot is framed from the chest down. The giant is never described, never turns toward camera and never speaks. The jacket, the needle, the spool of thread, the measuring tape and the worktable stay consistent in size, shape, position and appearance throughout. The tear exists until it is stitched, and once stitched it stays closed. 【Event Script】 Stage 1: A wide low shot along the surface of the worktable. The character walks steadily across the table toward the jacket spread out ahead of them, the coiled measuring tape carried over one shoulder. At the end, they are partway across the table and the jacket fills the ground ahead. Stage 2: They reach the edge of the jacket and stop at the foot of the fabric, tipping their head back to look up as the lapel rises above them. At the end, they are standing still at the base of the cloth, looking up. Stage 3: They climb onto the wool and cross to the failed seam, then crouch beside the open tear and lay a hand on the frayed edge, tracing its length. At the end, they are crouched at the tear with the split clearly open in front of them. Stage 4: They cross to the needle lying on the table and stand at its eye, raising the end of the thread with both arms and feeding it through the open loop. At the end, the thread is through the eye and they hold the doubled end. Stage 5: They set the needle into the wool and haul it through, braced back against the weight, the thread drawing after it in a long line. The cloth dimples and releases around each pass. At the end, the first stitches are laid and the needle is clear of the fabric again. Stage 6: A closer frame travels along the seam as the work continues, the tear closing behind a run of neat even stitches. They kneel at the last stitch and draw it tight. At the end, the tear is fully closed and the seam is clean. Stage 7: They step back onto the jacket, hands on hips, and look along the finished seam. An enormous hand descends into frame, closes on the cloth and lifts it, and they ride the rising fabric, catching their balance. At the end, the jacket is off the table and they are holding on. Stage 8: The camera pulls smoothly and steadily back into a wide shot. A giant stands wearing the repaired jacket, framed from the chest down. The character sits in the breast pocket with the needle held upright beside them, looking out. The shot holds on that final state. Throughout, movement is unhurried and physically believable, with correct anatomy and natural human-like motion. Objects have real weight — the needle is heavy, the cloth resists and settles, the thread has slack. Dust turns slowly in the raking light. Camera moves are smooth and cinematic with stable framing and consistent perspective; there are no rapid cuts, no camera jumps, no morphing, warping or artifacts, and no impossible physics. 【Sound】 <the soft scuff of small footsteps on polished wood>, <cloth shifting and settling under weight>, <the fine rasp of thread drawn through wool>, <a single quiet tap as the needle is set down>, <a distant clock in another room>, <the low hum of a warm afternoon workshop>. No one speaks, no narration, no music, and no subtitles appear on screen. The character keeps their mouth naturally closed throughout. 【Maintain Consistency】 Keep the character's identity, face, hair, clothing, proportions, colours and accessories exactly as defined by @image1 in every shot, with no drift or reinterpretation. Keep the workshop, worktable, jacket, needle, thread, measuring tape and lighting consistent with @image2, and hold the scale relationship between the character and every object fixed throughout. Keep the giant's face and head out of frame in every shot they appear in. Keep the tear closed once it is stitched. Keep the shot order, framing and camera geography matching @image3. Do not generate any text, titles, captions, subtitles, logos or watermarks.
Create a 15-second ultra-realistic cinematic 9:16 video of @product being built as a luxury miniature countryside villa, then revealed at golden hour. Keep @product locked to the reference: same silhouette, proportions, materials and colours once the form appears. One object, never duplicated. Two maker hands only — no faces, no other people. 0–3s — empty plot. A miniature countryside diorama: lush grass, olive trees, natural rocks, rustic wooden fencing, rolling farmland and distant hills under an orange-pink sunset. The footprint of @product is marked in the earth. Soft cinematic aerial, shallow depth of field. 3–7s — villa rises. Two hands, macro, miniature tools. Beige natural-stone walls of @product are laid, cement spread, a rustic terracotta tiled roof set, wooden doors and windows fitted, a covered outdoor dining patio added. Warm interior lights begin to glow. No cuts that break the same miniature. 7–11s — pool and water. Hands tile a crystal-clear turquoise swimming pool in front of @product: stone paving, wide steps, a dramatic infinity-style waterfall edge. Water fills; the cascade starts to flow with realistic liquid physics and photorealistic reflections. 11–15s — aerial pull-back. Smooth cinematic aerial pull-back revealing completed @product: glowing interiors, outdoor furniture, lanterns, olive trees, flowers, waterfall edge, the whole miniature property in golden-hour farmland. Hold the hero. No on-screen text. Sound: tiny masonry taps, water fill, waterfall, evening insects, a soft wind. No dialogue. No titles, no subtitles, no watermark, no logos that are not already on @product. Duration: exactly 15 seconds. Aspect ratio: 9:16.
10-second 9:16 colorful giant-product city film. @image1 is the product: photocopy that tall black matte tumbler with handle, black lid, stainless rim and base. Not a soda can. No invented label. No on-screen text. 0.0–1.33s: Extreme macro of the matte black wall and handle in bright sunlight. Cut wider into cyan sky. 1.33–2.67s: The same tumbler floats among chunky pink and yellow voxel blocks and miniature cargo stacks. Geometric glyphs orbit outside the product. Energetic scale shift. 2.67–4.0s: Dramatic low city angle. Giant tumbler towers over stacked cargo. Hard sunlight, cyan sky. Full geometry intact, handle visible. 4.0–5.33s: Top-down miniature cargo yard. Oversized tumbler among pink, yellow, cyan stacks. 5.33–8.0s: Open-sky. Blocks assemble around the tumbler. One brief water splash ~6.4–7.2s without deforming it. Splash settles by 8.0s. 8.0–10.0s: Low-angle oversized hero against cyan sky. Tumbler fully visible, larger than surrounding structures. Hold. No captions. AUDIO: original electronic beat, dry clicks, one splash accent, no vocals.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 10 seconds. AUDIO: Generate synchronized natural audio for this scene — ambient light rain outside, a rising whoosh as time freezes, near-silence during the freeze, then a sharp splash and a delighted laugh as time resumes. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. Photocopy this person. PHOTOREAL LIVE-ACTION — HIGHEST PRIORITY AFTER IDENTITY: Real camera, real tiny human, real jacket pocket. Not CGI, not a figurine, not a doll, not a sticker, not a composite. Match grain, colour and lighting between the person, the droplet and the fabric so the shot is one photograph. SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Exact face, facial features, eye shape, skin tone, build and gender presentation. Dark-brown hair in a high sculpted bun with a braid and two gold-and-green butterfly clips, smoky eye makeup, large geometric emerald-and-silver drop earrings, teal velvet jacket with gold-and-green floral embroidery. Do not beautify, restyle or alter identity. Only scale changes. SHE IS INSIDE THE POCKET BAG — NOT ON IT: The subject is a few centimetres tall and sits DOWN INSIDE the jacket-pocket cavity. Fabric walls surround her. She stands and moves on the inner fabric floor, not perched on the outside of the pocket like a brooch. Her legs and hips sink into the cloth; folds wrap her; she casts a contact shadow; she occludes the weave behind her. POCKET READ FROM FRAME 1: This is a real JACKET POCKET on a different full-sized adult’s coat (never the subject’s face on the wearer). A first-time viewer must instantly recognise a clothing pocket — not a fabric cave, not a bag, not a tunnel of cloth, not a person glued onto the pocket front. Camera sits just outside the garment at chest/hip height, looking down and IN through the pocket opening. The opening is angled slightly upward toward grey drizzle sky so daylight and rain can enter. In every frame keep visible at once: - the coat: outer fabric, seam, stitching, welt or flap, pocket-edge hem - the pocket as a 3D pouch sewn onto the jacket, with a clear opening - the tiny subject inside that pouch, a few centimetres tall, pocket bag huge around them Do not crop so tight that only fabric walls remain. If the pocket silhouette would leave the frame, pull the camera back. Outdoors, light drizzle, pale daylight shaft through the opening, dust and mist in the beam. Same light on her skin, the droplet and the fabric. Pores, flyaways, fabric weave. cameraFixed unset. CONTINUOUS MOTION: Real visible physical motion for almost the entire 10 seconds — no calm establishing pause longer than 1 second except the designed freeze below. 0-1s HOOK — already mid-action: pale daylight cuts into the pocket. She is already reacting, shielding her eyes, off-balance, down inside the pouch. Pocket opening, coat stitching and sky visible. 1-3s: One giant raindrop breaks through the opening and falls straight toward her in real time, growing larger. WOW — TIME FREEZE + ORBIT: The instant the drop is inches from her, everything FREEZES. Spherical droplet hangs, refracting daylight, faintly reflecting her silhouette. She is frozen mid-flinch, hair and jacket caught in motion, mist particles suspended. Camera does a slow cinematic 180-degree orbit around the frozen droplet and subject, shallow depth of field, still showing the pocket opening and coat. Freeze holds about 2 seconds. TIME RESUMES at full speed — droplet bursts past her in a splash, fine spray outward. 6-8s: Half-drenched, she laughs, wipes her face, notices a small puddle forming on the fabric floor beside her, her reflection in it. Still down inside the pouch; pocket hem and coat stitch in frame. 8-10s: She scoops a little puddle in cupped hands, watches it catch the light, looks straight into camera and grins triumphantly, still dripping, still inside the cavity. SCALE: Pocket stitching and fabric weave, raindrop surface tension and puddle reflection read at true scale next to the tiny subject. No on-screen text, no extra watermark, no subtitles, no second copy of her face on the wearer, no gym, no vanity lights from the still.
MODEL: Seedance 2.5 — use this model specifically. Do NOT use MiniMax or any other model for this generation. FORMAT: Vertical video, 9:16 aspect ratio. DURATION: 12 seconds. AUDIO: Generate synchronized natural audio for this scene — soft ambient fabric rustling, a held breath as the subject hides, a loud metallic jingle when the keys are pulled out, then a relieved exhale and a soft triumphant exclamation at the end. No music, no voiceover, no narration. ASSET MAP: @image1 — the subject. SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Exact face, facial features, hairstyle, skin tone, build and gender presentation stay unchanged — only their scale changes. Do not beautify, restyle or alter identity. Photocopy @image1. Face, hair, skin, and the olive-green quarter-zip over the white tee from the still stay. POCKET READ — MUST BE OBVIOUS FROM FRAME 1: This is a real jeans FRONT POCKET on blue denim jeans worn at the hip of a different, full-sized adult (never the subject’s face). A first-time viewer must instantly recognise it as a clothing pocket — not a fabric cave, not a bag, not a tunnel of cloth. Camera sits OUTSIDE the garment, slightly above hip height, looking down and in through the pocket opening. In every frame keep visible at once: - the jeans: waistband, a belt loop, bartack stitching, rivet, denim weave - the pocket as a 3D pouch sewn onto the jeans, with a clear U-shaped opening and a folded pocket-edge hem - the tiny subject inside that pouch, a few centimetres tall Do not crop so tight that only fabric walls remain. If the pocket silhouette would leave the frame, pull the camera back. CONTINUOUS MOTION REQUIRED: Real, visible physical motion for almost the entire 12 seconds — a clear sequence of distinct actions (settle, alert, hide, reaction, climb, look out), never one static pose held for seconds at a time. SCENE: The subject, shrunk to a few centimeters tall, is inside that jeans pocket. A folded receipt and a set of keys rest on the fabric floor nearby. The pocket opening is above them. Indoors, soft natural window light from one side. The wearer of the jeans is a different full-sized adult — never give them the subject’s face. 0-2s: Wide enough to read the jeans pocket on the hip. Tiny subject sits comfortably inside the pocket bag, arranging the folded receipt like a small mat beneath them, settled and at ease. Waistband and belt loop still in frame. 2-4s: Bright light suddenly widens above as the pocket flap is pulled open — they look up sharply, alert, rising to their feet. The U-shaped pocket opening stays readable. 4-6s: A pair of giant fingers (the wearer’s, not the subject) reach down into the pocket, searching side to side. The subject scrambles and presses themselves flat into a fold of fabric in the shadowed corner, perfectly still, visibly holding their breath. 6-8s: The fingers close around the keys resting just beside the subject’s hiding spot and lift them straight up and out of the pocket — the keys jingle loudly, a gust of motion ruffling the subject’s hair and clothing as they pass close by. Keys and fingers are enormous next to them. 8-10s: The fingers withdraw and the flap falls closed behind them, light dimming back down. The subject peeks out from the fold, relieved, then breaks into motion, climbing quickly up the inner fabric wall toward the pocket’s rim. Pocket silhouette still in frame. 10-12s: They reach the rim and perch on the edge, silhouetted against the bright opening, looking out at the towering room beyond, then turn back toward camera and give a small triumphant fist pump, grinning. Waistband still readable. SCALE REALISM: Stitching, denim weave, keys and fingers are true size next to the tiny subject. Match lighting so it reads as one photograph, not a composite. No on-screen text, no watermark, no subtitles.
SUBJECT LOCK — HIGHEST PRIORITY: Preserve the male subject in @image1 exactly — photocopy his exact face, facial features, hairstyle, skin tone, build and gender presentation throughout. Do not beautify, restyle, de-age or alter his identity in any way. Only his scale changes. No other people share his face. @image1 is the subject. Keep the charcoal three-piece suit, white dress shirt, burgundy patterned tie, gold ring and metal wristwatch exactly as shown. Only size changes: he is a few centimetres tall. CLIMB LOCK — CRITICAL, THIS WAS WRONG LAST TAKE: He must climb like a normal person on a vertical cliff. The OPEN PALM is the GROUND he stands on. The SIDE of the giant hand / wrist is the WALL he climbs. Chest faces that wall. Both hands reach UP to skin folds above him. Feet push on the palm, then find footholds on the side of the hand. Head up, looking toward the watch. Face and front of the body readable to camera — he is on the CAMERA-FACING side of the hand, never the far/hidden/opposite side. FORBIDDEN (do not do any of these): - sitting, crouching or scooting on his backside along the inner wrist - treating the inner wrist as a floor he crawls across - climbing the far side of the wrist so we only see his back - hanging from the buckle before he has climbed - a watch worn on the inner/palm side of the wrist — the giant watch sits on the OUTER/DORSAL wrist, the normal place a watch is worn CONTINUOUS MOTION REQUIRED: Real, visible physical motion for almost the entire 12 seconds — climb, leap, swing, landing. Never one static pose. SCENE: The subject, a few centimetres tall, starts standing in the centre of an open human palm — a different, full-sized adult hand, not the subject’s own. Indoors, soft natural window light from one side. Photoreal. One continuous real photograph, not a composite. The giant wrist wears a fabric-strap watch on the OUTER wrist, high above him at the top of the climb. 0-2s: He stands upright on the centre of the palm (feet on palm skin — this is the ground). He looks UP the vertical side of the hand toward the watch on the outer wrist. He starts climbing that wall like a rock climber: both hands grab skin folds near the wrist ABOVE his head, feet push off the palm, body upright against the wall, face visible. Real effortful climbing. Not a sit. Not a scoot. 2-4s: Still chest-to-wall, he reaches the top of the outer wrist, braces with his feet on the side of the hand, and leaps UP to grab the edge of the fabric watch strap with both hands. He swings slightly on impact, then steadies. He arrived by climbing the wall, not by crawling along the inner wrist. 4-6s: The giant hand slowly rotates at the wrist, lifting him into the air still gripping the strap. He holds on tightly as the rotation carries him through a wide arc, hair and clothing blown by the motion, expression shifting from tension to excitement. 6-8s: At the peak of the arc, he lets go deliberately, pushing off with his legs into a controlled fall back toward the palm below. 8-10s: He lands on the palm in a tuck-and-roll to absorb the impact, rolling once across the skin before catching himself on one knee. 10-12s: He rises, dusts off his hands, plants them on his hips, and breaks into a confident grin, pumping one fist in the air in triumph. SCALE REALISM: Hand skin, wrist folds, strap stitching and fabric weave read at true full size relative to the tiny subject. Match lighting and colour grade. Camera wide enough that the climb up the wall, the leap to the strap, the arc and the landing all stay in frame. Never open on the attached studio still or a grey headshot. No on-screen text, no watermark, no subtitles. AUDIO: Synchronized natural audio — soft ambient room tone, fabric rustling as he climbs, a faint metallic clink from the watch strap, breathing and effort under the climb, a triumphant exhale and laugh at the landing. No music, no voiceover, no narration. 9:16. TOTAL DURATION: 12 SECONDS.
SUBJECT LOCK — HIGHEST PRIORITY: Preserve the subject in @image1 exactly. Their exact face, facial features, hairstyle, skin tone, build and gender presentation stay unchanged — only their scale changes. Do not beautify, restyle or alter identity in any way. Use they / their / them only. @image1 is the subject. Photocopy their face, wavy shoulder-length brown hair, dark-green waffle-knit sweater and gold pendant necklace. Only their size changes: they are a few centimetres tall. Full body is that same person at miniature scale. SCENE: The subject, shrunk to a few centimetres tall, stands on an open human palm — a different, full-sized adult hand, not the subject’s own. Photoreal close-up of that palm: natural skin, enormous creases and fingerprints, fingers that dwarf them. Warm indoor light matching the green knit. Shallow depth of field. One continuous real photograph, not a composite. No cafe interior, no bookshelves, no window street, no other faces — only the palm that holds them. CONTINUOUS MOTION REQUIRED: Real, visible physical motion for almost the entire duration — not a static pose held for seconds. The hand, the subject, or both is always moving. SCALE REALISM: The palm’s skin texture, lines and fingers must read at true full size relative to the tiny subject — palm lines and fingerprint texture look enormous next to them. Match lighting and colour grade between the subject and the hand so the shot reads as one real photograph, not a composite. Never open on the attached still, a cafe portrait, or a studio headshot. No on-screen text, no watermark, no subtitles. 9:16. TOTAL DURATION: 8 SECONDS. 0.0–1.0s OPEN PALM. Already on the giant open palm. The tiny subject stands in the centre, full body readable, a few centimetres tall. Brief calm beat. Palm lines look enormous next to their shoes. Fingers dwarf them. 1.0–4.0s FINGERS CLOSE. The giant hand slowly begins to close — fingers curling upward and inward around the subject like walls rising on every side. The subject notices immediately, looks up at the fingers closing in, takes a startled step back and reaches out to brace a hand against the nearest finger for balance as the palm tilts slightly during the motion. 4.0–6.0s PUSH. The fingers stop just short of fully closing, leaving the subject in a half-enclosed space with light filtering through the gaps between fingers. The subject deliberately pushes with both hands against the inside of the nearest giant finger — and the finger visibly flexes and gives a little in response to the push, a real physical reaction, not a static wall. 6.0–8.0s OPEN. The hand slowly begins to open back up. The subject relaxes, breaks into a small, satisfied, amused smile, still steadying themselves with one hand on the finger as light floods back in. End on the palm, not a cutaway.