Video & audio

Change locations inside a phone video with AI

Prepare the source assets, follow the supplied generation workflow, and inspect the result frame by frame before exporting.

Adapted from TUT-031, dated 2026-06-23.

What you'll make

A checked final video that follows the source workflow and is ready for export.

Change locations inside a phone video with AI

Who this is for

Creators, marketers, founders, and small teams producing short AI-assisted videos from source media.

Step by step

Run the workflow

  1. 01

    The full workflow behind how I made this exact video

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Follow the shown setup, keep the source order unchanged, and run one controlled test.
    Source example for change locations inside a phone video with ai.
    Use this visual to confirm the expected setup or result.
    Source example for change locations inside a phone video with ai.
    Use this visual to confirm the expected setup or result.
    Source example for change locations inside a phone video with ai.
    Use this visual to confirm the expected setup or result.
    Source example for change locations inside a phone video with ai.
    Use this visual to confirm the expected setup or result.
    Source example for change locations inside a phone video with ai.
    Use this visual to confirm the expected setup or result.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  2. 02

    Prepare the source material

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Record videos of yourself acting the scenes out
    Source example for the full workflow behind how i made this exact video.
    Use this visual to confirm the expected setup or result.
    Source example for the full workflow behind how i made this exact video.
    Use this visual to confirm the expected setup or result.
    Source example for the full workflow behind how i made this exact video.
    Use this visual to confirm the expected setup or result.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  3. 03

    Generate the source images

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Screen shot the keyframes from your recorded footage and drop the reference images into ChatGPT with my prompt (Next Page).
    2. These are the screenshots from my video
    Prompt setup used for prepare the source material.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  4. 04

    Overall Preview

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Maintain the same chat for all the image generation that belongs to the same scenes (Next page for page by page breakdown):
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  5. 05

    Generate the next reference image 1

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Copy the prompt below and replace any project-specific subject, product, location, or format details.
    2. Attach the source assets in the same order described by the prompt, then run one controlled test.
    Working asset
    Generate the next reference image prompt
    Do the image generation all inside the same chat thread inside your ChatGPT.
    
    Generate an image of changing the location of this man to be at an American roadside diner from the late 1970s. Remove the selfie stick. He is sitting at his table looking at his phone, waiting for his pancake. And this is a pov of the waiter holding the pancake walking towards him. Do not change the man face, do not change the framing of the camera.

    Prompt setup used for overall preview.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  6. 06

    Generate the next reference image 2

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Copy the prompt below and replace any project-specific subject, product, location, or format details.
    2. Attach the source assets in the same order described by the prompt, then run one controlled test.
    Working asset
    Generate the next reference image prompt
    Do the image generation all inside the same chat thread inside your ChatGPT.
    
    Generate an image of changing the location of this man to be at an American roadside diner from the late 1970s. Remove the selfie stick. He is sitting at his table putting down his phone and in front of him is the plate of pancake that he got served. Do not change the man face, do not change the framing of the camera.

    Prompt setup used for generate the next reference image.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  7. 07

    Generate the next reference image 3

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Copy the prompt below and replace any project-specific subject, product, location, or format details.
    2. Attach the source assets in the same order described by the prompt, then run one controlled test.
    Working asset
    Generate the next reference image prompt
    Do the image generation all inside the same chat thread inside your ChatGPT.
    
    Generate an image of changing the location of this man to be at an American roadside diner from the late 1970s. The camera framing is a profile shot from the left. He is sitting at his table, reaching a utensil with his left hand and ready to eat the pancake that he got served. Do not change the man's face, do not change the framing of the camera.

    Prompt setup used for generate the next reference image.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  8. 08

    Generate the next reference image 4

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Copy the prompt below and replace any project-specific subject, product, location, or format details.
    2. Attach the source assets in the same order described by the prompt, then run one controlled test.
    Working asset
    Generate the next reference image prompt
    Do the image generation all inside the same chat thread inside your ChatGPT.
    
    Generate an image of changing the location of this man to be at an American roadside diner from the late 1970s. The camera framing is a profile shot from the left. He is sitting at his table, taking a bite of the pancake that he was served. Do not change the man's face, do not change the framing of the camera.

    Prompt setup used for generate the next reference image.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  9. 09

    Generate the next reference image 5

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Copy the prompt below and replace any project-specific subject, product, location, or format details.
    2. Attach the source assets in the same order described by the prompt, then run one controlled test.
    Working asset
    Generate the next reference image prompt
    Do the image generation all inside the same chat thread inside your ChatGPT.
    
    Generate an image of changing the location of this man to be at an American roadside diner from the late 1970s. Remove the selfie stick; his left hand is holding a utensil instead, and he is sitting up straight instead of leaning forward. The camera ultra wide framing, frontal angle, slightly tilted. He is sitting at his table, chewing on the pancake, the pancake was already sliced as per previous image. Do not change the man's face, do not change the framing of the camera.

    Prompt setup used for generate the next reference image.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  10. 10

    Generate the next reference image 6

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Copy the prompt below and replace any project-specific subject, product, location, or format details.
    2. Attach the source assets in the same order described by the prompt, then run one controlled test.
    Working asset
    Generate the next reference image prompt
    Do the image generation all inside the same chat thread inside your ChatGPT.
    
    Generate an image of changing the location of this man to be at an American roadside diner from the late 1970s. His both hand are on his head in shock because the pancake is so good. The camera ultra wide framing but is close up framing on his face, at a frontal angle. Do not change the man's face, do not change the framing of the camera.

    Prompt setup used for generate the next reference image.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  11. 11

    Generate the next reference image 7

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Copy the prompt below and replace any project-specific subject, product, location, or format details.
    2. Attach the source assets in the same order described by the prompt, then run one controlled test.
    Working asset
    Generate the next reference image prompt
    Do the image generation all inside the same chat thread inside your ChatGPT.
    
    Generate an image of changing the location of this man to be at an American roadside diner from the late 1970s. Remove the selfie stick; The camera side angle from slightly higher framing. He is sitting at his table with the pancake already sliced as per the previous image. He opened his mouth as if he were shouting. Do not change the man's face, do not change the framing of the camera.

    Prompt setup used for generate the next reference image.
    Check the prompt field, reference order, and visible settings.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  12. 12

    Open the video editor

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. With your assets ready, it's time to edit your footages!
    2. I am using Higgsfield for my Seedance 2.0:
    Source example for generate the next reference image.
    Use this visual to confirm the expected setup or result.
    Expected result for open the video editor.
    Use this source result as the visual comparison point for the step.
    Expected result for open the video editor.
    Use this source result as the visual comparison point for the step.
    Expected result for generate a transformed scene.
    Use this source result as the visual comparison point for the step.
    Expected result for generate a transformed scene.
    Use this source result as the visual comparison point for the step.
    Expected result for generate a transformed scene.
    Use this source result as the visual comparison point for the step.
    Expected result for generate a transformed scene.
    Use this source result as the visual comparison point for the step.
    Expected result for generate a transformed scene.
    Use this source result as the visual comparison point for the step.
    Expected result for generate a transformed scene.
    Use this source result as the visual comparison point for the step.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  13. 13

    Set the edit rhythm

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Raw AI footage is your base. Post-production is where it becomes cinematic.
    2. Temporal Remapping (Speed Ramps):
    3. Use speed curves to create rhythm and emphasize key moments.
    Source example for generate a transformed scene.
    Use this visual to confirm the expected setup or result.
    Source example for generate a transformed scene.
    Use this visual to confirm the expected setup or result.
    Source example for generate a transformed scene.
    Use this visual to confirm the expected setup or result.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  14. 14

    Grade the colour

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Visual Enhancement: Glow
    2. Apply a "Soft Glow" or "Edge Glow" effect to simulate high-end lens diffusion and luxury aesthetics.
    3. Color Grading: The Final Look
    4. Final color adjustments allow you to refine the mood beyond the AI's base
    Expected result for set the edit rhythm.
    Use this source result as the visual comparison point for the step.
    Expected result for set the edit rhythm.
    Use this source result as the visual comparison point for the step.
    Expected result for set the edit rhythm.
    Use this source result as the visual comparison point for the step.
    Expected result for set the edit rhythm.
    Use this source result as the visual comparison point for the step.
    Expected result for set the edit rhythm.
    Use this source result as the visual comparison point for the step.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
  15. 15

    Polish and export the final edit

    Use the source inputs to complete this stage and produce a result that can be checked before continuing.

    1. Polish is what makes it feel expensive. Fine-tuning is what maximizes impact and gives your commercial a premium finish.
    Source example for grade the colour.
    Use this visual to confirm the expected setup or result.
    Common failure modes
    • The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
    • Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.

Before you use it

Final checks

  • Check the full result for changed people, products, text, timing, or layout.
  • Confirm the tool, model, controls, and cost before generating.
  • Keep private or unreleased material out of third-party tools.

Share your work

Made it? Show me.

Post your finished result on Instagram and tag @terencesia. I'll share a few standout results with the community.

Tag @terencesia on Instagram (opens in a new tab)

Next step

Run one low-cost test, record the first visible failure, and revise only the instruction or source asset tied to that problem.

Back to the tutorial library

Sources

What this guide relies on