Jump to main content

Your audio, our set

Hotel Lobby AI With Your Own Song

Hotel Lobby AI with your own song keeps the orange set and the single mic but changes the words. Upload two portraits plus a short clip of your track, a voice memo or a toast, and the lead mouths it while the backup reacts.

  • Use 2 to 15 s of audio
  • Slide to your best part
  • Audio is never billed
  • Pairs with your own choreography
1

Your two faces

A single face per picture. Left leads, right backs up.

2

Set

3

Take and sound

Price: 6 credits. Done in roughly 5 to 15 min; if it fails, the credits come back. See prices

Under the hood

How your song drives the duet

Think of a take as three inputs. The portraits decide who is on the set. A reference performance decides the room, the camera and the choreography. A reference track decides what the front voice is lip-syncing. Unless you change it, that track is the Hotel Lobby audio.

Upload your own file and only the track changes. The lead performs your words into the mic while the sidekick bounces and reacts, still on the orange wall. Drag the slider to choose where your clip begins; the browser trims it to the take length, 15 seconds at most, before anything is sent.

You can stack this with your own dance video too: the clip sets the moves, your audio sets the mouth. Only upload audio you are allowed to use.

Session plan

Making a Hotel Lobby video to your own track

  1. Illustration: two separate portrait photos, an audio file.
    01

    Cast two faces

    One clear portrait each. The left face is the front voice who performs your audio; the right face is the sidekick.

  2. Illustration: an audio file, a waveform with trim handles, headphones.
    02

    Drop in the audio

    Under Soundtrack & moves, choose Your song and add the file. If it is longer than the take, drag to the section you want to hear.

  3. Illustration: a video preview with a play button, an audio waveform, a download arrow.
    03

    Run the take, save the MP4

    6 credits for a standard 12-second take, whichever audio you use. Watch it and download; if the take fails, the credits are returned.

Casting ideas

Audio people bring to the set

Illustration: a voice recording on a phone, a birthday cake.
01

A birthday shout-out

Record ten seconds of a silly birthday verse on your phone, then cast the birthday girl and yourself to perform it.

Illustration: a vinyl record, an audio file, a speaker.
02

Your band’s new hook

Tease a release in the trend’s format: the chorus as the audio, two band members as the portraits.

Illustration: conversation bubbles, a voice recording on a phone.
03

A group-chat legend

That voice memo everyone still quotes, delivered like a COLORS session. Short, clearly spoken memos sync best.

Illustration: a notebook of rap ideas, a hanging microphone, an audio waveform.
04

Bars you wrote yourself

Write a verse roasting your roommate, record it over any beat, and let the front voice deliver it.

From your file to the finished take

  1. You choose the file. MP3 uploads as is; WAV, M4A, AAC and OGG get converted in your browser first.
  2. You drag to the part you want. The browser keeps exactly the take length (5 to 15 seconds) from that point and drops everything else.
  3. Your trimmed audio stands in for the Hotel Lobby track. The set, the camera, the choreography and the price stay as they were.
  4. Wan 3.0 generates the front voice lip-syncing your audio at the mic while the sidekick reacts, with your audio as the soundtrack. Watch it before posting; no artist’s voice is cloned.

Limits worth knowing

WorksDoes not work
The front voice performing your song, rap or memoBoth faces singing your upload (AI rap with Both rap covers that)
Up to 15 seconds, the size of a TikTok or Reels clipA full-length music video
Following your recording’s words and timingCloning a famous voice, or a perfect copy of your file
Wan 3.0, on the Hotel Lobby set or with your own dance clipCleaning up noisy recordings; clear vocals sync best

Would rather have a verse written for you? The AI rap generator writes an original rap on your subject with a Seedance model.

Our own takes

Hotel Lobby Video test renders

Two sample renders, with model, length and source shown under each. They demonstrate the workflow, not a promise about how often a take succeeds.

Made by us

Wan 3.0 · 480p · 10 s · 16:9

  • Pipeline test render, made outside the public website.
  • Sound: the Hotel Lobby song.
  • Compressed 854 × 480 copy for the web. The settings listed refer to the full-size take.
Made by us

Wan 3.0 · 720p · 12 s · 16:9

  • Owner test-account render from Oct 3, 2026 on an earlier site version.
  • Sound: the Hotel Lobby song.
  • Compressed 854 × 480 copy for the web. The settings listed refer to the full-size take.

Around the web

Hotel Lobby Clips by Other Creators

These clips show the format. Other people made them with other tools, and each one is credited and linked to its post. They are not output from Hotel Lobby Video.

Audio questions

Your own song: questions

Need a hand? Write to us or check the prices.

Which audio files can I upload?

MP3, WAV, M4A (what most phones save voice memos as), AAC or OGG. MP3 goes up untouched; the others are converted in your browser. The take uses 2 to 15 seconds, starting wherever you set the slider.

I recorded a voice memo for my dad’s retirement. Can my brother and I perform it, from my phone, without paying extra for the audio?

Yes. Upload the memo under Your song, add a portrait of each of you, and create. Audio never adds to the price: on Wan 3.0 at 480p it is 3, 4, 5, 6 or 7 credits for 5, 8, 10, 12 or 15 seconds.

Who sings it, and what does the other person do?

The left face mouths your audio at the mic. The right face points, laughs and dances on the beat, the way the original routine goes.

Can both of us sing my upload, trading lines?

No. Only the front voice lip-syncs an uploaded track. For two people trading lines, use AI rap with Both rap on a Seedance model, or make two takes with the portraits swapped and edit them together: two people, one song.

Can I upload the actual “Hotel Lobby” song?

Only if you hold the rights to that recording. For a post with the official sound, add the licensed version inside TikTok, Reels or CapCut when you publish.

Is there a free way to use my own song?

No. Every take is paid in credits, and your own audio costs the same as the stock track: 6 credits for 12 seconds. The smallest pack is 10 credits for $9.99, and failed takes are refunded.

I have no song. Can the AI write one?

Yes. Choose a Seedance model under Video, set the sound to AI rap and type what it should be about (up to 120 characters). The model writes an original verse and beat; it ignores any upload and is instructed not to copy an existing song.

My track is four minutes long. Will that upload?

Yes. The file can be any length; your browser cuts out just the 2 to 15 seconds you select, so the upload stays small.

How do I get the cleanest lip-sync?

Use a clear vocal with little background noise, and start the selection on a word, not on silence. Rapping and plain speech both work.

Your track, our set

Let the duet perform your song

Two portraits and one short audio clip. The same price as the stock track.