Start with a clear visual idea
Two individual portraits become the starting point for a night-time lobby scene: two performers, reflective stone, warm lamps and a camera that moves closer. The visual idea is a composed entrance with alternating attention between the pair, rather than a crowded concert.
A place with depth
The lobby supplies a broad floor, tall interior lines and reflections. These background cues establish a location without competing with the two faces.
Warm light, dark surroundings
Gold practical lighting and deeper green shadows define the intended palette. Skin should remain readable instead of being swallowed by the darker interior.
A paired performance
The intended camera language moves from a shared view toward individual attention. Small head and shoulder movements suit a short clip better than a complex dance request.
Three ways to describe the scene
Choose a single example as a starting brief in a prompt-enabled tool. Replace any subject description with a reference you have permission to use. Avoid stacking all three examples into one request.
The shared entrance
Make a vertical fictional music-video scene from these two individual portraits. Put the two subjects in a spacious hotel entrance at night, with warm lamps and softly reflected floor light. Keep both faces recognizable, outfits consistent and movement restrained. Bring the camera closer while preserving room around the pair.
This example prioritizes the paired scene framing. It avoids asking for several conflicting camera moves in a short sequence.
Portrait-led attention
Create a two-person lobby performance with gentle changes of framing. Begin with both subjects visible, then give each person a brief portrait moment. Use soft amber highlights against a dark interior. Preserve the reference faces and clothing; keep hands below the face and avoid extreme camera angles.
Naming what stays stable—faces, clothing and readable framing—is more actionable than repeating quality adjectives.
A calmer editorial version
Reimagine the pair as performers in an elegant hotel lobby. Use a quiet camera approach, natural posture and a minimal background. Let the lighting and reflections create the atmosphere. Avoid a busy crowd, large gestures, captions and brand logos; compose for a vertical phone screen.
A restrained alternative can describe the same setting without introducing choreography, a crowd or a specific copyrighted track.
What to make explicit
Describe the pair before the décor
Make the subject count and two separate portrait references explicit. Marble, haze and lamps are supporting details; they should not distract from stable faces.
Choose one camera intention
A gentle approach plus a change of portrait framing is easier to describe coherently than an orbit, crane move and rapid zoom in the same brief clip.
Separate the look from the song
Trading attention between two performers suggests rap-video staging. It does not specify lyrics, a licensed recording or accurate lip synchronization.
How these examples relate to this studio
The generator on this site takes one person in each of two separate portraits and applies a fixed Two-Person Hotel Lobby scene recipe. You do not need to type a prompt to use it, and pasting an example does not change its settings.
This scene defaults to 10 seconds in a 9:16 frame. You can choose a 10- or 15-second length in the generator, which shows current live availability and the selected credit ceiling before submission. Camera motion, expressions, scene details and any audio are generated, so an example is an intention rather than an exact output guarantee.
A rap-style visual does not guarantee a particular commercial song, lyrics, choreography or lip synchronization. Do not request a real person’s endorsement or claim that synthetic performance footage is an actual recording.
Questions before you start
Can I edit the prompt in this generator?
No. The current studio provides one fixed scene recipe, with one portrait for a solo scene or two separate portraits for a pair. These public examples are learning material, not a custom-prompt setting.
Are these the exact internal generation prompts?
No. They are separately written public examples that explain the effect’s visual choices. The private recipe, provider settings and billing controls are kept on the server.
Will each person sing a different verse?
The look suggests alternating performer attention. It does not guarantee words, a particular song, different recorded verses or synchronized mouth movement.
READY FOR YOUR REFERENCE?
Start with portraits you can use.
Preview each photo locally, then check live availability and the credit cost before generating.
Open Two-Person Hotel Lobby