Feedback
AI Ad Video Example
Loading...
Migos AI Video Generator
Give the Migos AI Video Generator two photos you have rights to, and it builds a vertical 9:16 duet — one mic, one orange booth.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Veo3.1
Create Stunning Videos with Veo3.1
FLUX 3 Video Generator

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator
Bring Your Own Cast Into a Shared Spotlight
This video-tools page converts a pair of photos you have permission to use into a two-person orange-booth performance, so you can ride the trend with faces you actually control.
- The Classic Orange-Booth StagingCompose a 9:16 frame from two separate cleared photos: one performer stands full-length on the left, the second mirrors on the right, all against an unbroken burnt-orange wall with a black microphone suspended at the center.
- Distinct Identities, Trading VersesEach face has to stay recognizable and locked to its own side. Let one performer open with natural gestures while the partner reacts, then hand the lead across without ever swapping positions.
- Inspiration Without Rights HeadachesThe trend is a springboard, not a permit. Cast fictional adults, consenting friends, or pets, and pair them with audio you recorded or licensed — never celebrity likenesses or copyrighted recordings.
Three Steps to Your Orange-Booth Duet
From a pair of reference photos to a finished two-person performance in a terracotta studio, this walkthrough covers every stage of the build.
What Keeps a Two-Person Clip Believable
Identity, timing, staging, and rights are all steered through the prompt — and that control is what separates a convincing duet from a muddled one.
Keeping Two Faces Apart
Label each subject LEFT or RIGHT, leave breathing room around the microphone so hands and torsos never fuse, and simplify motion before layering on style if the two identities start to blur.
Verse Trading on Cue
Write the handoff into the prompt: one voice holds the verse while the partner listens and answers, then the lead flips at the midpoint with sides untouched, using brief nods and small hand beats instead of constant motion.
Locked Orange-Studio Frame
Hold the shot to a matte-orange wall and floor with no visible seams, soft frontal light, one black microphone dead center, a wide locked angle, and both performers' feet inside the frame.
Casting You Have the Right to Use
Build the duo from fictional adults, willing friends, pets, or another original pairing, set it to audio you recorded or licensed, and never present generated celebrity footage as genuine.
Direct the Duet From the Prompt
You decide who opens, how the partner answers, and whether the energy stays low-key or builds — enough leeway to bend the format toward your own friendship, pet, comedy, or creator concept.
Fix One Flaw at a Time
Revise only the line that governs the problem — merged faces, colliding hands, or a drifting camera — so each new version is far easier to judge than a full restart.
Orange-Booth Duet: Common Questions
Answers on reference photos, rights, side swaps, and the rendering model behind the two-person orange-studio look.
Which visual cues signal the orange-booth duet style?
Three elements give it away — a burnt-orange backdrop, one microphone suspended between two performers, and a turn-taking exchange. The look traces back to Quavo and Takeoff's “Hotel Lobby” set on A COLORS SHOW.
Are two reference photos really necessary?
Two are strongly recommended. Telling the performers apart and pinning each one to a fixed side is the entire effect, so choose cleared shots with similar lighting and clearly visible faces.
May I use celebrity images or the original track?
Stay with material you have permission for — likenesses, images, clips, tracks, and vocals included — and never let a generated celebrity clip pass as real. Using your own subjects with audio you wrote or licensed keeps you on solid ground.
Why do the two faces keep blending into one?
Input quality and placement are usually the cause: dark, cluttered, or unassigned references leave the model with nothing to anchor on. Keep one subject per image, spell out LEFT and RIGHT in the prompt, and reduce overlapping gestures.
Which rendering model should I choose?
Seedance 2.5 is the default on this page because it handles multi-reference input and long, detailed direction well. Which models you can reach, and how many credits they cost, varies with your account and region.
How can I prevent side swaps and mirrored movement?
Keep each performer anchored to a single side from start to finish, spell out the turn-taking order step by step in the prompt, and let reactions stay small and personal instead of copying one another.
Make Your Own Orange-Booth Duet
Choose two subjects you have clearance to feature, describe how each one stands and responds, and let Seedance 2.5 render the finished two-person clip.
