Hotel Lobby AI for Couples: Make a Personalized Rap Video
Create a personalized 10-second couple rap video from one shared photo or two portraits, with scene choices, anniversary details, generated verse, and native audio.
The short answer
For a couple video, use one shared photo when both partners are clearly visible or two separate portraits when those give better identity references. Choose Hotel Lobby, Rooftop Night, Warehouse Session, or Neon Garage, then select Anniversary or use a general detail. MiniMax M3 writes a short original verse from the information you provide, and MiniMax H3 generates the final 10-second vertical performance with synchronized beat and rap vocals.
A couple version does not need a separate romantic video engine. The useful personalization comes from three inputs the current workbench already understands: who the two people are, what visual scene fits them, and what specific relationship detail should shape the short verse.
That makes the format useful for a playful anniversary post, a relationship joke, a long-distance surprise, or simply a more personal version of the Hotel Lobby trend. It is still a short rap performance, not a long love-story video, and that boundary is worth keeping clear.
Best for
- ✓ Partners who want a short funny or stylish personalized duo clip
- ✓ Anniversary posts or relationship posts where one specific detail matters
- ✓ Couples with either a strong shared photo or clear individual portraits
Not for
- × A long romantic slideshow or multi-scene anniversary film
- × A custom full-length love song
- × Manual editing of the generated verse before video generation
Choose one shared photo or two portraits based on clarity
A good couple photo can be a strong one-photo input because both people are already visible together. Use it when both faces are large enough to recognize and neither person is hiding the other. If the favorite couple photo is dark, distant, or heavily overlapping, the safer route is two separate portraits.
Do not choose two portraits simply because the tool supports them. Choose them when they genuinely provide cleaner identity information than the shared image.
Pick the visual mood before you write the personalization detail
Hotel Lobby keeps the recognizable orange performance setup and its motion reference. Rooftop Night gives the pair a city-at-night music-video feel. Warehouse Session is more industrial and performance-focused, while Neon Garage gives a brighter late-night visual style.
Choosing the scene first helps you keep the text detail focused. The detail does not need to describe the camera or background because the scene preset already handles that. Use the text field for the relationship itself.
Use the Anniversary option as a starting point, not a restriction
Selecting Anniversary tells the MiniMax M3 lyric step what kind of moment the verse should celebrate. You can then add names, years together, a location, a shared habit, or another short fact. If the clip is not for an anniversary, leave the occasion general and describe what you actually want the short verse to recognize.
The generated verse is intentionally compact because the final video is only 10 seconds. A detail such as “Nina and Alex, five years together, survived three apartments and one very stubborn dog” gives the model more useful material than a generic request for a romantic rap.
What happens after you click generate
Step 1
The system validates the photo mode and scene
One-photo mode expects one image containing both people; two-photo mode expects exactly two image references.
Step 2
MiniMax M3 writes the short verse
The occasion and personal detail are converted into a compact original English rap verse designed for the 10-second clip.
Step 3
MiniMax H3 generates the performance
The source image references, selected scene, and generated verse are sent into H3 with instructions for synchronized rap vocals, beat, motion, and two-person identity preservation.
Step 4
The result is stored in your account history
Successful videos can be reopened and downloaded later. Technical failures or canceled generation tasks return the generation Credits.
Use the free preview as a scene check, not a lyric preview
Before sign-in, the site offers one free static scene preview. That preview can help you compare how the pair looks in Hotel Lobby versus Rooftop Night or another preset before you commit to the full video.
It does not generate the MiniMax M3 verse and it does not contain H3 motion or audio. If the main question is “do we look good together in this visual setup?” the preview can answer it. If the main question is “does the rap line feel personal?” that is part of the full generation.
Current product boundaries
| Feature | Current guided workflow |
|---|---|
| Photo input | One photo together or two separate portraits |
| Scenes | Hotel Lobby, Rooftop Night, Warehouse Session, Neon Garage |
| Personalization | Occasion plus up to 320 characters of detail |
| Verse | Generated automatically by MiniMax M3 |
| Video model | MiniMax H3 |
| Output | 10-second 9:16 video at 768p with native audio |
| Cost | 100 Credits per current scene |
Try the focused workflow
Preview your duo in a rap-video scene first
Use one photo together or two separate portraits, choose one of four scenes, and create one free static preview before sign-in. For the full video, add an occasion or personal detail so MiniMax M3 can write the short verse before MiniMax H3 generates the 10-second rap performance.
Frequently asked questions
Can I use one wedding photo with both of us?
Yes, if both faces are clearly visible and large enough to identify. If the image is a wide group wedding photo, crop or use two individual portraits instead.
Can I add our names to the rap?
Yes. Add the names and one useful relationship detail in the personalization field. MiniMax M3 decides how to incorporate that information into the short original verse.
Can I write the exact lyrics myself?
The current guided Hotel Lobby AI workbench is built around automatic verse generation from your details. It does not expose a manual lyric editor before the H3 video generation.
