What you'll make
A checked final video that follows the source workflow and is ready for export.
Create a cinematic AI ad from concept to final edit
Who this is for
Creators, marketers, founders, and small teams producing short AI-assisted videos from source media.
Step by step
Run the workflow
- 01
Workflow breaks down on how I made this exact video
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Follow the shown setup, keep the source order unchanged, and run one controlled test.

Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 02
Prepare clear reference images of the character featured in your video
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Choose Your Character
- Prepare clear reference images of the character featured in your video. Use multiple angles and expressions to help AI maintain identity consistency across scenes.

Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 03
Part 1, Script Generation
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Your Personal Director
- Describe your rough scene idea and use ChatGPT to expand it into structured scenes with camera direction, emotional tone, and dialogue cues. 1

Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 04
Part 2, World Building (Image Prompts)
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Your Personal Director
- Next, convert your script into structured image prompts.
- Convert your script into detailed image prompts that define environment, lighting, character pose, and camera perspective. These prompts generate first-frame anchors that guide video motion and consistency.

Check the prompt field, reference order, and visible settings. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 05
Open the image generator
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Click here to open Google Flow (opens in a new tab)Direct link supplied in the source. Source page 9.
- With your assets ready, it's time to turn your phone photos into commercial-looking ad frames.
- Go to Google Flow and select "Create with Flow."
- Select + New project

Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 06
Upload the source image
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- You may receive free credits (varies by account).
- Upload your source image.
- Copy the selected prompt from ChatGPT and paste it into Google Flow.

Check the selected files and their order before continuing. 
Check the selected files and their order before continuing. 
Check the selected files and their order before continuing. 
Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 07
Generate the video 1
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Transform your keyframe into motion using Kling AI. Precise prompts improve motion realism and help preserve your intended commercial look.
- Open Kling AI → click Generate → choose Video Generation (Video 3.0). 1

Check the prompt field, reference order, and visible settings. 
Check the prompt field, reference order, and visible settings. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 08
Generate the video 2
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Upload your Phase 2 generated image as the First Frame, then paste the scene-specific video prompt. 2

Check the selected files and their order before continuing. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 09
Scene Output: 3
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Repeat for Remaining Scenes
- You can use Kling AI's Bind Elements to maintain character consistency by attaching a reference image to your subject. However, this feature is less suitable for crowded scenes, as it may affect the facial appearance of surrounding people.
- *Best used for single-subject or close-up scenes.

Use this source result as the visual comparison point for the step. 
Use this source result as the visual comparison point for the step. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 10
Generate the final scene
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- To create additional scenes, repeat the workflow starting from Phase 1 (Part 1):
- 1.Tell ChatGPT what you want for the next scene 2.Generate the video script (scene direction +
- dialogue) 3.Generate the image description/prompt for that
- scene 4.Use the generated image as the First Frame in Kling
- 5.Paste the corresponding scene video prompt and
- Important: Not every scene will work on the first attempt. If motion or realism breaks, refine the script or prompt before regenerating to minimize unnecessary credit usage.

Check the prompt field, reference order, and visible settings. 
Check the prompt field, reference order, and visible settings. 
Check the prompt field, reference order, and visible settings. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 11
OPERATIONAL NOTES
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Click here to open Kling AI (opens in a new tab)Direct link supplied in the source. The Workflow Lab may earn a commission. Source page 17.
- Quality control: This workflow is optimized for Professional Mode. Standard/Free modes may produce lower resolution or unstable motion (jitter).
- Get the Tools Here
- Let's continue to the next phase
Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 12
Set the edit rhythm
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Raw AI footage is your base. Post-production is where it becomes cinematic.
- Temporal Remapping (Speed Ramps):
- Use speed curves to create rhythm and emphasize key moments.

Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. 
Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 13
Grade the colour
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Visual Enhancement: Glow
- Apply a "Soft Glow" or "Edge Glow" effect to simulate high-end lens diffusion and luxury aesthetics.
- Color Grading: The Final Look
- Final color adjustments allow you to refine the mood beyond the AI's base

Use this source result as the visual comparison point for the step. 
Use this source result as the visual comparison point for the step. 
Use this source result as the visual comparison point for the step. 
Use this source result as the visual comparison point for the step. 
Use this source result as the visual comparison point for the step. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
- 14
Polish and export the final edit
Use the source inputs to complete this stage and produce a result that can be checked before continuing.
- Click here to open CapCut (opens in a new tab)Direct link supplied in the source. The Workflow Lab may earn a commission. Source page 20.
- Polish is what makes it feel expensive. Fine-tuning is what maximizes impact and gives your commercial a premium finish.

Use this visual to confirm the expected setup or result. Common failure modes- The generated scene changes the person, camera, room, product, or timing beyond the requested effect.
- Fast motion hides face, hand, clothing, text, or continuity errors that become visible when the clip is paused.
Before you use it
Final checks
- Check the full result for changed people, products, text, timing, or layout.
- Confirm the tool, model, controls, and cost before generating.
- Keep private or unreleased material out of third-party tools.
Share your work
Made it? Show me.
Post your finished result on Instagram and tag @terencesia. I'll share a few standout results with the community.
Tag @terencesia on Instagram (opens in a new tab)Next step
Run one low-cost test, record the first visible failure, and revise only the instruction or source asset tied to that problem.
Back to the tutorial librarySources
What this guide relies on
- Google Flow (opens in a new tab)labs.google. Reviewed 2026-08-18.
- Kling AI (opens in a new tab)klingaiaffiliate.pxf.io. Reviewed 2026-08-18.
- CapCut (opens in a new tab)capcutaffiliateprogram.pxf.io. Reviewed 2026-08-18.