Prompt
Style: Authentic iPhone UGC, 9:16 vertical, static front-facing selfie shot, natural indoor lighting. Single continuous shot — zero camera cuts, zero transitions, zero visual effects. The entire video must be indistinguishable from a real person filming themselves. SCENE 1 [00:00–00:03] CAM: Medium close-up, static, front-facing selfie framing with the slightest natural handheld drift. A young man (mid-20s, pleasant but unremarkable face, light stubble, slightly messy hair, wearing a plain crewneck sweater) sits in what appears to be his bedroom — neutral background, a shelf with a few books, a plant, soft natural daylight coming from a window to his left. The framing is casual, slightly off-center — the way someone would prop their phone to talk to a friend. Nothing about the image, the person, or the setting raises any suspicion. He looks directly into the lens with a natural, easy expression — the kind of half-smile that precedes an embarrassing story you've told before and can now laugh about. He speaks with a conversational, confessional tone — not performing, not scripted, just telling a story. [CHARACTER SPEAKS ON CAMERA - NOT a voiceover] DIALOGUE (natural, warm, slightly self-deprecating): "Il mio primo giorno di lavoro è stato un disastro totale." LIP-SYNC MANDATORY: Character opens mouth, lips move with every word. Audio comes FROM the character on screen, not from a narrator. SCENE 2 [00:03–00:08] CAM: Same exact shot. No change in framing, lighting, or position. Continuous single take. ACTION: He shifts into the story with genuine micro-expressions — each detail triggers a specific facial reaction that sells the authenticity. When he mentions the clothes, his eyes glance down briefly at himself with a wince. When he mentions the coffee machine, he shakes his head slightly, lips pressed together in remembered embarrassment. When he says "Lei," his eyebrows lift and he makes a small grimace — the face of someone reliving a social miscalculation. After the Lei/tu detail, a brief pause. His gaze drops to the lower left — the universal micro-expression of someone remembering something lonely. His voice softens slightly for the final detail. [CHARACTER SPEAKS ON CAMERA - NOT a voiceover] DIALOGUE (continuous, natural pacing with micro-pauses between details): "Sono arrivato vestito troppo elegante. Non capivo la macchinetta del caffè. Ho dato del Lei al capo... mentre tutti si davano del tu." LIP-SYNC MANDATORY: Character opens mouth, lips move with every word. Audio comes FROM the character on screen, not from a narrator. A beat. A half-smile forms — not happy, more like the rueful acceptance of someone who has processed this memory enough to find it slightly funny. [CHARACTER SPEAKS ON CAMERA - NOT a voiceover] DIALOGUE (quieter, slower): "E ho pranzato da solo alla mensa." LIP-SYNC MANDATORY: Character opens mouth, lips move with every word. Audio comes FROM the character on screen, not from a narrator. SCENE 3 [00:08–00:12] CAM: Same exact shot. Same room. Same light. Same person. Same position. Absolutely nothing changes visually. This is critical — there must be zero visual signal that a reveal is coming. No glitch, no distortion, no color shift, no music cue, no zoom. The continuity must be seamless. ACTION: His tone shifts subtly — lighter, almost amused, as if he's about to deliver a punchline he's been holding back. The half-smile stays. His eyes remain locked on the camera. He tilts his head very slightly. [CHARACTER SPEAKS ON CAMERA - NOT a voiceover] DIALOGUE (lighter tone, conversational, a beat before each sentence): "Ma la verità è che io non ho mai lavorato un giorno in vita mia." LIP-SYNC MANDATORY: Character opens mouth, lips move with every word. Audio comes FROM the character on screen, not from a narrator. A pause. One second of silence. His expression doesn't change. He holds eye contact with the lens. Then, with the same casual tone, same warmth, same face — as if stating something entirely ordinary: [CHARACTER SPEAKS ON CAMERA - NOT a voiceover] DIALOGUE (same tone, no dramatic shift, no robotic delivery): "Perché io non sono reale." LIP-SYNC MANDATORY: Character opens mouth, lips move with every word. Audio comes FROM the character on screen, not from a narrator. FINALE [00:12–00:15] CAM: Same shot. Same person. Same room. He remains still, looking directly into the camera. His expression is neutral — not smiling, not sad, not threatening. Just present. The same face that told a relatable, empathetic story about a first day at work now holds your gaze in silence. He doesn't blink for the first two seconds. The room is completely quiet. No sound. No ambient noise. No music. Nothing. Three seconds of a face that was human thirty seconds ago and is now something else — same pixels, same light, same everything, but the context has permanently shifted. The silence forces the viewer to sit with the realization. He breathes once, softly, naturally — and the video ends. DIALOGUE: None. Complete silence. Technical specs: 15 seconds total, 9:16 vertical, 4K, 24fps. CRITICAL: Absolutely no text overlays, no captions, no on-screen copy of any kind. Absolutely no visual effects, no glitch effects, no distortion, no AI-aesthetic cues. The entire point is that the video is visually indistinguishable from a real human UGC selfie video from beginning to end. LIGHTING: Soft natural daylight from window, consistent throughout — no changes, no shifts, no dramatic shadows. Warm, natural skin tones. Shallow depth of field softly blurring the background. Authentic iPhone front-camera aesthetic — slight wide-angle lens distortion at edges, natural skin texture, no beauty filter. Audio: natural room tone throughout Scenes 1-3 (subtle air, faint ambient hum), transitioning to absolute silence in the Finale — the removal of room tone itself creates the discomfort. No music at any point. The dialogue delivery must be entirely natural, conversational Italian — no robotic cadence, no unnatural pauses, no AI voice artifacts. The micro-expressions must be specific and human: embarrassment, self-deprecation, loneliness, amusement, and finally neutral stillness. Single continuous shot — zero cuts from start to finish. No filters, no color grading shifts, no slow motion. SPEAK IN THE SAME LANGUAGE AS THE DIALOGUE TEXT. If dialogues are written in English, speak English. If in Italian, speak Italian. Never switch to a different language. MANDATORY AUDIO (HIGHEST PRIORITY): This video MUST have audible spoken dialogue — a silent video is FAILED. The character MUST speak out loud with clear, loud voice audio. NO silent/mute output. Audio is REQUIRED. ALL audio MUST come from the on-screen character's mouth with visible lip sync. NO voiceover, NO narrator, NO off-screen voice. SILENCE RULE: After dialogue ends, the character MUST stay completely SILENT. No filler words, no random sounds. Just visual action until the scene ends. PHYSICAL REALISM: All actions must be physically possible for a real human. NO walking next to moving vehicles, NO reaching through objects, NO floating/teleporting. STATIC OBJECTS (parked cars, furniture, buildings) MUST stay completely still — ZERO movement. Only the character moves. AUDIO MIX: Voice must be LOUD and CLEAR, like a close-up mic recording. Voice is the PRIMARY audio. Background/ambient sounds barely audible or absent. NO wind, hum, crackling, or digital artifacts. The character MUST open their mouth and physically speak. This is NOT a voiceover. NOT narration. The audio source is the character on screen. MANDATORY: This video MUST have an audio track with spoken dialogue. Silent output = failed generation. NO text overlay, NO subtitles, NO captions on screen.
Style: Authentic iPhone UGC, 9:16 vertical, static front-facing selfie shot, natural indoor lighting. Single continuo...
Instellingen
Maak iets soortgelijks
De generator hieronder selecteert vooraf het model dat voor dit record wordt gebruikt, zodat je het idee sneller kunt remixeren.