The Hotel Lobby AI video generator: upload 2 photos, pick the COLORS template, hit generate, download in HD. No green screen, no reference clip, no login to preview.
Your video may be selected and published to the public Explore feed to inspire other creators.
Examples
The ai hotel lobby format is a character slot machine: any two faces can share the orange booth. These are the combos the internet keeps remaking — find yours, then shoot the same style of portraits.
Cross-species duo: deadpan blinking on the beat, the dog leans in at the drop.
Make this duoUnder the hood
The hotel lobby generator is a template-driven image-to-video pipeline: two portraits in, one orange-booth performance out. Here is every step, every parameter, and where model versions actually differ.
Drop a clear, front-facing portrait into slot one. Good lighting and an unobstructed face give the model the landmarks it needs for likeness — no sunglasses, no heavy filters.
Add the second portrait to slot two. This performer stands on the right of the two-person medium shot and usually answers the first verse.
The template presets the solid orange backdrop, the single hanging vintage microphone, the tight two-shot framing, and vertical 9:16 output — the exact stage that makes the trend recognizable.
6 seconds is the classic trend length; 10–15 seconds fits a full verse. Choose resolution (480p drafts fast, 768p+ for HD downloads) and the model tier that balances speed and fidelity.
The hotel lobby ai video generator animates both photos into the orange-booth performance — lip-sync, head-nods, and gestures — usually within one to three minutes. No login wall before preview.
Download the finished vertical clip, then lay the Hotel Lobby audio back over it in CapCut, InShot, or the TikTok editor (mute the generated audio). Post with #hotellobby and #genjutsu.
| Parameter | What to set | Notes |
|---|---|---|
| Photos required | 2 portraits (one per performer) | Front-facing, one face per photo; a single group photo works but weakens likeness. |
| Stage | Solid orange backdrop (COLORS show) | Preset by the template — no green screen, no set dressing. |
| Microphone | Single hanging vintage/condenser mic | Centered between the two performers; the visual anchor of the trend. |
| Framing | Tight two-person medium shot | Both performers share the frame from the chest up, 9:16 vertical. |
| Duration | 6s / 10s / 15s (model-dependent 5–15s) | 6 seconds is the classic loop; longer durations cost more credits. |
| Resolution | 480p draft → 768p–1080p HD | Draft at 480p to check likeness, re-render winners in HD. |
| Aspect ratio | 9:16 vertical | Image-to-video can inherit the input photo's ratio — crop to 9:16 for TikTok/Reels/Shorts. |
| Audio | Model-dependent native audio | Most creators mute generated audio and sync the original Hotel Lobby song in post. |
The hotel lobby ai generator exposes more than one model tier, and they trade speed against fidelity. Pick by what the render is for:
Runs 5–15s renders at up to 768p with the strongest facial-identity transfer. Best when the faces must actually look like your two photos. Slower per render.
Sub-3-second-class renders at a slightly lower fidelity ceiling — perfect for testing prompts and photo pairs cheaply before committing credits to an HD re-render.
Some model versions generate synchronized audio with the clip. For the trend itself, mute it and add the original Hotel Lobby track — that is what viewers expect to hear.
Prompts
These are the exact prompt bodies behind the most popular hotel lobby ai video outputs. Copy one, paste it into the prompt box above, and swap in your duo. Many more scenarios live in the hotel lobby prompt library.
Two best friends performing as a rap duo on a minimalist stage: solid orange backdrop, single hanging vintage microphone, tight two-person medium shot, 9:16 vertical. Performer on the left gestures while rapping; performer on the right nods to the beat, both smiling, natural lip-sync to the beat, casual streetwear, warm studio lighting.Two fictional characters reimagined as a hip-hop duo performing live: minimal orange stage set, single hanging condenser mic centered between them, medium two-shot, 9:16. The character on the left delivers the first verse with confidence; the character on the right reacts and answers on the second verse, cinematic lighting, shallow depth of field.A couple recreated as a rap duo in an iconic orange-booth music performance: one central hanging microphone, plain orange stage, two-person medium shot, vertical 9:16. They take turns at the mic, playful eye contact between verses, synced head-nodding, soft key light from the front, subtle film grain.Two pets "performing" as a rap duo: orange minimalist stage, one hanging vintage microphone between them, two-shot framing, vertical video. Gentle head-bobbing and blinking on the beat, one leans toward the mic at the drop, adorable but deadpan, warm studio light, crisp fur detail.Fixes
Four issues cause almost every bad render. Here is the fix for each — try these before spending credits on an HD re-render.
Fix — Use a higher-resolution, front-facing photo with even lighting and nothing covering the face (hair, sunglasses, masks). Crop tight to the face so the model spends its attention on identity, not background. Avoid group photos and heavy beauty filters — both smear facial landmarks.
Fix — Generate at 6 seconds first: shorter clips keep head-nods and lip-sync tight. Put the beat cues in the prompt ("nods on the beat", "gestures on the verse") and mute the generated audio in post, syncing the real Hotel Lobby track to the render — matching cuts to the beat is what sells it.
Fix — Hands are the hardest body part for every video model. Prompt around it: "hands relaxed at their sides", "one hand on the mic stand", or frame it as a medium shot where hands sit at the edge of frame. Re-render rather than fight a bad seed — failed renders are not charged here.
Fix — Queues spike when the trend trends. Draft at 480p and 6 seconds on the fast tier to iterate quickly, then run the winning photo pair at higher resolution. Renders that fail in a queue are refunded automatically — you never pay for a video you can't download.
Every angle of the trend, one page each: