
Good starting point
- One person, with the face and hairstyle clearly visible.
- Waist-up or a wider crop that shows the outfit.
- Even light and enough detail around the eyes.
- Natural proportions without a strong beauty filter.
15-second party preview. Generation is not yet available.Explore the 15-second party scene and its lead role. Preview only; generation is not yet available.
A new lead in a familiar party: warm chandeliers, a metal cup and gestures straight into the camera.
Burning Bridges AI videos replace the lead in a crowded indoor performance. The recognizable details are the handheld camera, pointing and waving, the metal cup, and the guests around the performer. This template uses the first 15 seconds of the landscape party reference shown here.
See the party referenceThis is the 16:9 party excerpt, not a full 29-second edit or a 20-second vertical remake. Your photo defines the lead's face, hair, body and visible outfit; the reference defines the performance.
The useful idea is simple: change the lead without inventing a different scene.
The source centers on a performer in a black leather jacket holding a metal cup. Our fixed excerpt keeps the first 15 seconds and the original landscape frame.
Party reference and scene detailsA Reddit discussion shares a reference-video workflow for replacing the lead while preserving gestures, camera and light. The same thread includes a question about price and an unresolved content-restriction report.
Read the community workflow discussionUpload the person you want in the lead role. This version keeps the original crowd instead of assigning your photo to everyone.
This is the planned workflow. Generation stays closed until we can show a reviewed example.
Use one clear photo showing the face, hair and clothes you want. It replaces the foreground performer with the metal cup, not the crowd.
The clip is 15 seconds in 16:9. Choose 480P for 225 credits or 720P for 375 credits, and check the total before generating.
Watch the finished video, especially the cup, hands and quick head turns. Your result stays in History so you can return and download it.
Choose a photo that clearly defines one person and the outfit you want in the party.


The photos define the person. The video defines the performance.
The prompt uses your photo for the lead's identity and visible clothes, not just a new face on the source body.
Only the foreground lead is assigned to your photo. Background guests are instructed to stay unchanged.
The reference guides pointing, waving, cup movements and handheld framing. Fast motion or occlusion can still introduce differences.
The source audio is restored to the finished excerpt. This does not clone your voice or create a new song.
Both options use the same 15-second landscape party excerpt.
480P costs 225 credits. 720P costs 375 credits. The higher resolution changes output detail, not the scene or duration.
The Generate button shows the full credit cost. Buying a credit pack and spending credits on this clip are separate steps.
This template uses fal's Wan 3.0 Prime reference-to-video workflow. Generation quality varies, and a new attempt is a separate generation. Model reference-to-video documentation
Compare the workflows, not a promised winner. CapCut templates vary, and every AI result needs review.
Keep your photos and explore more character-replacement templates.
One photo of one person. This template replaces only the foreground lead with the metal cup. It does not offer background-person replacements.
It is the first 15 seconds of the landscape party reference shown on this page: a crowded room, warm chandeliers and a lead performer with a metal cup. It is not the full 29-second reference or a separate 20-second vertical template.
The prompt asks the model to take the lead's face, hair, skin, body proportions and visible outfit from your photo. It is a whole-person replacement workflow, but small accessories, lettering, hands or clothing may still change.
Not in this version. Your photo is used only for the foreground lead, and the prompt keeps the original background people. Crowded or overlapping faces can still drift, so review them in the result.
The output is 15 seconds in 16:9. It costs 225 credits at 480P or 375 credits at 720P. Check the final total on the Generate button before submitting.
The finished clip uses the reference excerpt's audio. It does not clone your voice, record you singing or compose a new song. Sharing the audio is subject to the rights and rules of the platform where you post it.
No. The template already includes a lead-only replacement prompt and a fixed motion reference. Upload the photo and choose the resolution.
No. The model is instructed to preserve the actions, camera, timing and scene, but exact reproduction is not guaranteed. The metal cup, hands crossing the face, dark light and fast turns are important moments to check.
A terminal generation failure returns the credits charged for that task. A completed video with visual imperfections is not the same as a failed task. Check your result and task status before starting another generation.
No. This is an independent AI character-replacement tool. The trend name describes the scene people are looking for, not an endorsement. Use photos, footage and audio that you have permission to transform and share.