The Hotel Lobby AI video generator turns two clear portraits into a coordinated 15-second performance. Person 1 stays on the left, Person 2 stays on the right, and both follow a fixed sequence of gestures beneath the suspended microphone.
Instead of asking you to write a prompt, Vidoro keeps the Hotel Lobby AI video generator focused on one recognizable template: the saturated orange studio, locked camera and original movement timing. That makes the setup simple and keeps every attempt comparable.
The process begins with identity replacement rather than free-form scene creation. Your first photo supplies the face and recognizable traits for the existing left performer. Your second photo does the same for the right performer. The template remains the source of truth for body position, standing pose, camera perspective, floor, lighting and microphone placement.
Headshots and half-body photos can still be useful. When a source image does not show the lower body, the system keeps the template pose and extends only the reliable visual information it can see. Face shape, key features and hairstyle take priority over inventing an entirely new outfit or body. This gives identity consistency a clearer target during the later movement sequence.
Your first generation is a free 720P preview with a small Vidoro watermark. If the cast and motion look right, one-time credits unlock watermark-free 720P or 1080P output.