Process

Tools

Techniques

  • vid2vid

  • img2vid

Generation data

COPY ALL

Prompt

External Generator
For the target video, at 0.00 seconds into the target video, <Picture 1> (from [Shot 1]) is fully referenced. integrated_multimodal_description: [Shot 1] Live-action, soft indoor portrait look, a chest-up close-up begins fully locked to <Picture 1>, fully preserving the woman’s appearance, clear rectangular glasses, side braid, pastel-blue sweatshirt, plain wall, and clean daylight color tone: her fingertip is already on the frame; she gives the glasses a small upward push, blinks, then smiles. Clothes stay fully on. The young woman (S1) inhales, then starts to sing: [Chinese] 眼镜后面偷看你一眼…… The camera pushes in with small amplitude at slow speed toward her eyes. [Shot 2] At 00:02.700, the camera cuts to a tighter close-up of her mouth and lenses: she sways her head a little, the glasses catching a soft highlight as she sings. (S1) continues: [Chinese] 心跳偷偷告诉我……喜欢你。 The handheld frame pans with small amplitude at slow speed from the glasses corner to her peach-tinted lips. [Shot 3] At 00:05.400, the shot eases back as she covers a shy giggle, pushes the glasses once more, then leans toward the lens to finish. (S1) sings, voice extra small: [Chinese] 今晚也要梦见你呀…… The camera holds, then settles with tiny amplitude at slow speed on her still-clothed close-up as the clip lands at about 8.00 seconds. No on-screen captions, titles, lyrics, or extra lettering. overall_soundscape: Quiet room hush and a faint sleeve rustle when she adjusts the glasses. Her close-mic singing sits dry and intimate, with a tiny breath before each phrase. non_diegetic_music: Soft music-box pop with sparse piano and a warm pad that gently follows her vocal, thickening slightly under each cut, fading on a held chime near 8.00 seconds.

Other metadata

cfgScale:1
steps:19
sampler:Euler a

Discussion