Two photos in · one viral video out

Hotel Lobby AI Video Generator

The Hotel Lobby AI video generator: upload 2 photos, pick the COLORS template, hit generate, download in HD. No green screen, no reference clip, no login to preview.

Required
PublicFeatured

Your video may be selected and published to the public Explore feed to inspire other creators.

Your video will appear here after you click Generate.

Examples

Hotel Lobby AI Video Examples by Character Duo

The ai hotel lobby format is a character slot machine: any two faces can share the orange booth. These are the combos the internet keeps remaking — find yours, then shoot the same style of portraits.

Drake & Best Friend

Soft-spoken opener, hype-man answers on the second verse.

Make this duo

Barack & Michelle Obama

Playful eye contact between verses, synced head-nodding.

Make this duo

Palmer Luckey & Founder

Tech-founder deadpan delivery with confident hand gestures.

Make this duo

Iron Man Solo Booth

One performer, arms wide on the hook — the template solo variant.

Make this duo

Two Best Friends

Streetwear duo: left gestures while rapping, right nods to the beat.

Make this duo

Wedding Couple

Newlyweds take turns at the single hanging mic, warm celebratory light.

Make this duo

Dog & Cat

Cross-species duo: deadpan blinking on the beat, the dog leans in at the drop.

Make this duo

Brothers

Family-portrait energy — older brother raps, younger brother clowns.

Make this duo

Goku & Naruto Class Duo

Anime rivals forced to share one mic under #genjutsu.

Make this duo

Boss & Intern

Office duo in business casual, competitive verse-trading energy.

Make this duo

Under the hood

How the Hotel Lobby AI Video Generator Works

The hotel lobby generator is a template-driven image-to-video pipeline: two portraits in, one orange-booth performance out. Here is every step, every parameter, and where model versions actually differ.

  1. 1

    Upload the first performer's photo

    Drop a clear, front-facing portrait into slot one. Good lighting and an unobstructed face give the model the landmarks it needs for likeness — no sunglasses, no heavy filters.

  2. 2

    Upload the second performer's photo

    Add the second portrait to slot two. This performer stands on the right of the two-person medium shot and usually answers the first verse.

  3. 3

    Pick the Hotel Lobby COLORS template

    The template presets the solid orange backdrop, the single hanging vintage microphone, the tight two-shot framing, and vertical 9:16 output — the exact stage that makes the trend recognizable.

  4. 4

    Set duration, resolution, and model

    6 seconds is the classic trend length; 10–15 seconds fits a full verse. Choose resolution (480p drafts fast, 768p+ for HD downloads) and the model tier that balances speed and fidelity.

  5. 5

    Hit Generate and watch the render

    The hotel lobby ai video generator animates both photos into the orange-booth performance — lip-sync, head-nods, and gestures — usually within one to three minutes. No login wall before preview.

  6. 6

    Download in HD and add the song

    Download the finished vertical clip, then lay the Hotel Lobby audio back over it in CapCut, InShot, or the TikTok editor (mute the generated audio). Post with #hotellobby and #genjutsu.

Generator parameters

ParameterWhat to setNotes
Photos required2 portraits (one per performer)Front-facing, one face per photo; a single group photo works but weakens likeness.
StageSolid orange backdrop (COLORS show)Preset by the template — no green screen, no set dressing.
MicrophoneSingle hanging vintage/condenser micCentered between the two performers; the visual anchor of the trend.
FramingTight two-person medium shotBoth performers share the frame from the chest up, 9:16 vertical.
Duration6s / 10s / 15s (model-dependent 5–15s)6 seconds is the classic loop; longer durations cost more credits.
Resolution480p draft → 768p–1080p HDDraft at 480p to check likeness, re-render winners in HD.
Aspect ratio9:16 verticalImage-to-video can inherit the input photo's ratio — crop to 9:16 for TikTok/Reels/Shorts.
AudioModel-dependent native audioMost creators mute generated audio and sync the original Hotel Lobby song in post.

Model version differences

The hotel lobby ai generator exposes more than one model tier, and they trade speed against fidelity. Pick by what the render is for:

Standard tier (best likeness)

Runs 5–15s renders at up to 768p with the strongest facial-identity transfer. Best when the faces must actually look like your two photos. Slower per render.

Fast/turbo tier (best for drafts)

Sub-3-second-class renders at a slightly lower fidelity ceiling — perfect for testing prompts and photo pairs cheaply before committing credits to an HD re-render.

Audio-capable tiers

Some model versions generate synchronized audio with the clip. For the trend itself, mute it and add the original Hotel Lobby track — that is what viewers expect to hear.

Prompts

Hotel Lobby AI Video Prompts (Copy & Paste)

These are the exact prompt bodies behind the most popular hotel lobby ai video outputs. Copy one, paste it into the prompt box above, and swap in your duo. Many more scenarios live in the hotel lobby prompt library.

Best friends — the classic hotel lobby ai video
Two best friends performing as a rap duo on a minimalist stage: solid orange backdrop, single hanging vintage microphone, tight two-person medium shot, 9:16 vertical. Performer on the left gestures while rapping; performer on the right nods to the beat, both smiling, natural lip-sync to the beat, casual streetwear, warm studio lighting.
Movie characters — cinematic hotel lobby generator prompt
Two fictional characters reimagined as a hip-hop duo performing live: minimal orange stage set, single hanging condenser mic centered between them, medium two-shot, 9:16. The character on the left delivers the first verse with confidence; the character on the right reacts and answers on the second verse, cinematic lighting, shallow depth of field.
Couple — playful duo variant
A couple recreated as a rap duo in an iconic orange-booth music performance: one central hanging microphone, plain orange stage, two-person medium shot, vertical 9:16. They take turns at the mic, playful eye contact between verses, synced head-nodding, soft key light from the front, subtle film grain.
Pets — deadpan pet duo
Two pets "performing" as a rap duo: orange minimalist stage, one hanging vintage microphone between them, two-shot framing, vertical video. Gentle head-bobbing and blinking on the beat, one leans toward the mic at the drop, adorable but deadpan, warm studio light, crisp fur detail.

Fixes

Hotel Lobby AI Video Troubleshooting

Four issues cause almost every bad render. Here is the fix for each — try these before spending credits on an HD re-render.

Problem: The face doesn't look alike

Fix — Use a higher-resolution, front-facing photo with even lighting and nothing covering the face (hair, sunglasses, masks). Crop tight to the face so the model spends its attention on identity, not background. Avoid group photos and heavy beauty filters — both smear facial landmarks.

Problem: The rhythm / lip-sync is off

Fix — Generate at 6 seconds first: shorter clips keep head-nods and lip-sync tight. Put the beat cues in the prompt ("nods on the beat", "gestures on the verse") and mute the generated audio in post, syncing the real Hotel Lobby track to the render — matching cuts to the beat is what sells it.

Problem: Hands or mic look broken

Fix — Hands are the hardest body part for every video model. Prompt around it: "hands relaxed at their sides", "one hand on the mic stand", or frame it as a medium shot where hands sit at the edge of frame. Re-render rather than fight a bad seed — failed renders are not charged here.

Problem: Generation is slow or queued

Fix — Queues spike when the trend trends. Draft at 480p and 6 seconds on the fast tier to iterate quickly, then run the winning photo pair at higher resolution. Renders that fail in a queue are refunded automatically — you never pay for a video you can't download.

Hotel Lobby AI Video Generator FAQ

It is an image-to-video pipeline tuned to one scene: you upload two portraits, and the model animates both people as a rap duo on the COLORS-style orange stage — single hanging mic, tight two-person medium shot, vertical 9:16. The template presets the stage, framing, and motion so you never write a full prompt from scratch.
One clear, front-facing portrait per performer, evenly lit, with nothing covering the face. Two separate photos beat one group photo because the model maps exactly one identity to each performer. Photos from the chest up at 1:1 or 4:5 crop cleanly into the vertical stage.
Your first generations are free to preview — no login required to see the result. Free renders are short vertical clips; HD downloads and longer durations use credits. See the free tier page for the exact current allowance, and remember failed renders are never charged.
Usually one to three minutes depending on queue load, duration, and model tier. Fast/turbo tiers draft a 6-second clip quickest; standard tiers take a little longer for stronger likeness. Progress is shown inline and the download unlocks the moment the render finishes.
No. The orange booth, the hanging microphone, and the lighting are generated by the AI. The only editing you might do is dropping the original song over the downloaded clip in a mobile editor — a 30-second job in CapCut or the TikTok app itself.
Use the fast/turbo tier for drafts and likeness checks, the standard tier for faces that must actually look like your photos, and 768p or higher for the HD download you post. Longer durations (10–15s) need more credits, so lock your photo pair with a cheap draft first.
Video models generate their own audio (or none), and licensing stops any generator from shipping the track inside the clip. The standard workflow: download the video, mute generated audio, add the song in your editor, and align the head-nods to the beat before posting.
Only upload photos you have the right to use. Celebrity-face memes are everywhere in this trend, but publicity rights vary by country and platform, and commercial use is a different matter entirely. The safe route is yourself, friends, family, pets, or characters you have licensed.
Yes to both. Pet duos are one of the most-loved variants — use the deadpan pet prompt. Solo performers work too: one portrait in the first slot and a solo-stage prompt (Iron Man-style solo booth) turns the duo format into a one-person performance.
Yes — it is the same hotel lobby trend format: the orange COLORS stage, the single mic, the medium two-shot. This hotel lobby generator just removes the friction: no reference clip, no character-sheet setup, no login wall before preview, and no charge for failed renders.

Keep exploring the hotel lobby trend

Every angle of the trend, one page each: