← AI Creator CourseAll programsHomeSearch
AI Creator Course·Full Video Walkthroughs·2:13:12

Section 6: Four Full Builds — a Faceless Hook in 20 Minutes, a '$300,000 Ad for Under $30', an Animated Music Video, and a Viral AI-Clone Action Sequence

Anthony Gallo Instructor - builds the faceless hook and the Dr. Squatch-style ad end to end; frames the clone lesson · Sam Chung Guest presenter - student musician who animated his own music video (ChatGPT + Nano Banana Pro + Hailuo + Canva + CapCut) after a traditional-animation attempt stalled on cost · Laurie Possible presenter of the AI-clone action sequence ('I'm Laurie for contentcreator.com'); may be an ASR misread

The short version

  1. A faceless hook in ~20 minutes: ChatGPT script iterated to under 500 characters -> ElevenLabs narration -> Suno instrumental ('no vocals... suspenseful'), generated twice -> Resolve: trim pauses, slide the music so its crescendo lands on the strongest line -> Seedance 2.0 multi-shot prompt with timestamps matched to the narration, its own audio deleted -> a too-short shot regenerated in Kling ('doesn't cost as many credits') -> SFX in two buckets, what you SEE and what you FEEL, lengthened and turned down. 'You cannot be broad with your AI prompting... if you are broad, it's likely your fault.'
  2. The ad ($300k look for under $30, Dr. Squatch style): ChatGPT outline; characters from a template (four identical vague prompts gave four different people; the template gave four consistent ones); a Google Doc storyboard with each still under its script line; Veo 3.1 for dialogue and first/last-frame transformations (soap swap, young-to-old dad in one second), Kling 2.6 for product and reaction shots and for the boy - Veo refuses anyone under 18; ElevenLabs V3 with bracketed director cues ('[depressed]', '[excited]') and Voice Changer to re-voice Veo's baked-in dialogue; a long Resolve edit.
  3. Consistency is edit chains, not regeneration: hover 'Edit' on a Nano Banana output to reuse it as the base; 'swap the face in image one with the face from image two, maintaining realism and matching the original lighting and expression'; 'keep everything in the scene exactly the same, but replace the person.' The clone pipeline: outfit -> environment -> face swap from multi-angle headshots, 4K, four variations, repeated per scene 'until I got it perfect' - 'the secret sauce' against slop.
  4. Motion prompts fail until you use the trade's words: 'kong vault', 'precision jump', 'safety roll', 'julienne', 'flambe' - 'AI is trained on professional footage that uses correct terminology.' Wan 2.2 Animate MOVE (not Replace) transfers real iPhone footage - head, arms, weight shifts, micro-expressions - onto the clone; a Kling transition bridges it into the animated sequence; zoom-in cuts hide the face seams; a compound clip carries one global speed ramp.
  5. Sam's music video: feed the old animatic MP4 to ChatGPT to extract characters and story (correcting its misreads), lock one Pixar-style master prompt, front-view character sheets, environments built frame by frame from the sketches with image-to-image blending, a Hailuo preset ('no exposure change, make video smooth'), never upscale, salvage 2-second sub-clips from bad shots, and run image gen, video gen and CapCut in parallel.

At a glance, three clicks deep

Skim here first: the closed row is the glance, open is the study card with the key points and timestamps, and the ↓ link drops to that concept's full write-up below.

01The faceless hook in 20 minutes: script -> voice -> music -> Seedance -> SFXIterate the script short;

Iterate the script short; narrate; score; sync by crescendo; one timestamped Seedance prompt; patch shots in Kling; SFX seen/felt; ~20 minutes.

Script iterated to <500 chars (p2198688192 0:04)

Suno: vocal-free suspense, generated twice (p2198688192 0:06)

Crescendo aligned to the strongest line (p2198688192 0:07)

Seedance 2.0 multi-shot with timestamps; delete its audio (p2198688192 0:10-0:12)

Kling for cheap regeneration (p2198688192 0:13)

SFX seen vs felt; longer; quieter (p2198688192 0:19-0:20)

'If you are broad... it's likely your fault' (p2198688192 0:17)

↓ Full write-up of this concept

02The '$300,000 ad for under $30': template characters, Google Doc storyboard, Veo + Kling, ElevenLabs V3Script -> template characters -> Doc storyboard -> Veo (dialogue, transformations) + Kling (product, reacti…

Script -> template characters -> Doc storyboard -> Veo (dialogue, transformations) + Kling (product, reactions, minors) -> ElevenLabs V3 cues + Voice Changer -> Resolve edit.

Vague vs template prompts: consistency demo (p2194234933 0:05)

Google Doc storyboard; Web Page export for full-res (p2194234933 0:08-0:19)

Veo 3.1: dialogue, tripod drift, refuses under-18 -> Kling 2.6 (p2194234933 0:22-0:25)

First/last frame for one-second transformations (p2194234933 0:23)

ElevenLabs V3 director cues; Voice Changer on baked dialogue (p2194234933 0:27-0:31)

Resolve layering so baked audio never covers narration (p2194234933 0:38)

↓ Full write-up of this concept

03Consistency comes from edit chains and face swaps, not regenerationEdit-in-place from approved outputs;

Edit-in-place from approved outputs; explicit swap prompts naming image one and image two; repeat the face swap per scene; blend references for undescribable detail.

'Edit' reloads a prior output as the base (p2194234933 0:11-0:12)

Standard face-swap prompt (p2194234933 0:12)

Clone pipeline: outfit -> environment -> face swap; 4K x4 (p2195581486 0:04-0:06)

Reapply per scene 'until I got it perfect' (p2195581486 0:08)

Image-to-image reference blending (p2194175296 0:19-0:20)

↓ Full write-up of this concept

04Motion prompts: start and end states, and the trade's vocabularyMotion prompt = precise start state + end state + the professional term for the move.

Motion prompt = precise start state + end state + the professional term for the move.

Vague motion prompts 'absolutely sucked' (p2195581486 0:09)

Start/end states plus domain terms (kong vault, flambe, safety roll) (p2195581486 0:09-0:10)

Kling 2.6 as go-to over Veo 3 for detail (p2195581486 0:09)

↓ Full write-up of this concept

05Wan 2.2 Animate 'Move': put your real performance on the cloneStoryboard -> iPhone base + headshots -> base clone -> scene stills -> Kling animation -> Wan 2.2 Move tran…

Storyboard -> iPhone base + headshots -> base clone -> scene stills -> Kling animation -> Wan 2.2 Move transfer -> Kling bridge -> masked edit.

Seven-step pipeline (p2195581486 0:02-0:17)

Move vs Replace distinction (p2195581486 0:13)

Zoom transitions mask face mismatches (p2195581486 0:15)

Compound clip for a global speed ramp (p2195581486 0:16)

↓ Full write-up of this concept

06Sam's animated music video: animatic in, master prompt, Hailuo preset, salvage the good two secondsExtract story from existing sketches with ChatGPT;

Extract story from existing sketches with ChatGPT; one master prompt; front-view sheets; Hailuo with a stability preset; salvage sub-clips; parallel pipelines; cut by feel.

Stack and no-upscale rule (p2194175296 0:07-0:08)

Animatic MP4 -> ChatGPT character/story extraction with corrections (p2194175296 0:09-0:13)

Master prompt template; front-view sheets; blended environments (p2194175296 0:14-0:22)

Hailuo preset; avoid end-frame stitching (p2194175296 0:24-0:29)

Salvage 2-second sub-clips; multiple failed shots (p2194175296 0:27-0:36)

CapCut beat markers 'wasn't that important' (p2194175296 0:36-0:37)

↓ Full write-up of this concept

07The edit is the human differentiator: Resolve, CapCut, Premiere habits from all four buildsResolve for assembly and audio discipline;

Resolve for assembly and audio discipline; zoom cuts and nested speed ramps to hide seams; plugin SFX; deep-dive in the editing section.

Resolve free, one-time pro fee (p2198688192 0:08)

Layer tracks so baked audio never covers narration (p2194234933 0:38)

Zoom transitions mask face mismatches (p2195581486 0:15)

Nest/compound clip for global speed ramp (p2195581486 0:16)

Beat markers optional (p2194175296 0:36-0:37)

↓ Full write-up of this concept

The concepts in full

01

The faceless hook in 20 minutes: script -> voice -> music -> Seedance -> SFX

how-to

Front-load everything on the first minute. If the hook holds, the rest gets watched.

Three pillars: idea, quality, packaging; AI enhances judgment, it does not replace it (Wendover Productions as the model). Script: ChatGPT iterated three times ('write a one minute hook' -> 'diving right into the moment' -> 'under 500 characters'). Narration: ElevenLabs TTS on PromptEdit. Music: Suno, 'no vocals, just a slow pulsing background track... suspenseful', two generations to choose from. Resolve: trim narration pauses, duplicate/reposition the music so the crescendo hits the most intense line, 6-frame crossfade at the music edit. Visuals: ChatGPT drafts a Seedance 2.0 prompt in 15-second segments with scene timestamps matching the voiceover; delete Seedance's baked-in audio and lay the clip over the narration; a too-short opening drone shot is screenshotted and regenerated longer in Kling; split and re-time. SFX from ElevenLabs' SFX tab in two buckets - seen (volcano, footsteps, cups rattling, dog) and felt (riser, heartbeat, impact) - with 'choose specific duration' checked, lengthened, volume down so they 'enhance... not distract.' Title card from the Content Creator Templates. Total ~20 minutes.

Do it in this order
Why it matters

A complete, timed recipe for the faceless-channel product Paul's clients keep asking about.

02

The '$300,000 ad for under $30': template characters, Google Doc storyboard, Veo + Kling, ElevenLabs V3

how-to

Four identical vague prompts gave four strangers. The template gave the same man four times. That is the whole difference between an ad and slop.

Recreating a Dr. Squatch-style spot (the brand's '$5 million to over $100 million in a year' is cited). Outline and script in ChatGPT. Casting: a character-prompt template on Higgsfield's Nano Banana Pro - the vague-vs-template comparison. Storyboard: paste the script into a Google Doc, drag only FINAL images under their lines (resize freely; quality is unaffected), later File > Download > Web Page to get a full-res image folder for Higgsfield. A scene-generation template (camera framing, subject/props, lighting/tone, background action, mood) is a free download. Video on Higgsfield: Veo 3.1 for dialogue and tone (but it drifts off a locked tripod and 'refuses to generate anyone appearing under 18' - the boy goes to Kling 2.6); Kling 2.6 for product and reaction shots, more detail retention, fewer credits; first/last frame on Veo for one-second magical transformations (weak soap to Dr. Squatch bar; young to aged dad). Voices: ElevenLabs Explore filtered by language/accent/age, '+' saves an Australian narrator and a deep dad voice; V3 model with bracketed director cues ('[depressed]', '[excited]'), 'way more expressive and human' than V2, two takes per request; Voice Changer re-voices Veo's baked-in dialogue to the saved voice, then frame-nudged to the waveform in Resolve. A long real-time Resolve assembly, layering tracks so a clip's own audio never covers the narration.

Do it in this order
Why it matters

This is the most complete paid-ad recipe in the KB, with the two model gotchas (tripod drift, under-18 refusal) that cost hours.

03

Consistency comes from edit chains and face swaps, not regeneration

Hover the image, click Edit, and the model starts from what you already approved.

Never regenerate from scratch: reuse an approved output as the base and layer changes. Product swap: 'Replace the weak soap with the bar of Dr. Squatch soap' with a real product photo uploaded. Face swap: 'swap the face in image one with the face from image two, maintaining realism and matching the original lighting and expression.' Before/after: 'keep everything in the scene exactly the same, but replace the person... with the person from the second image.' The clone pipeline formalizes it: outfit transform -> environment transform -> face swap from multi-angle, multi-expression headshots, at 4K with four variations per pass on fal.ai, the same swap prompt reapplied to every scene 'until I got it perfect' - 'the secret sauce to making this look professional instead of AI slop.' Sam blends references image-to-image (a rooftop plus a street) for details he cannot phrase.

Why it matters

This is the single practice that separates the finished pieces in this section from everything on the feed.

04

Motion prompts: start and end states, and the trade's vocabulary

'Roll on the floor' failed five times. 'Safety roll' - the parkour term - worked first time.

Early motion prompts 'absolutely sucked' until two changes: describe the exact start and end state of the motion, and use domain terminology - parkour ('kong vault', 'precision jump', 'cat leap', 'safety roll'), cooking ('julienne', 'saute', 'flambe') - because 'AI is trained on professional footage that uses correct terminology.' Kling 2.6 is the presenter's go-to 'including over Google Veo 3' for character detail and multi-object scenes; Cinematic scene stills came from ChatGPT-authored image-to-image prompts per storyboard bullet.

Why it matters

A five-word rule that fixes most 'the model did something else' complaints.

05

Wan 2.2 Animate 'Move': put your real performance on the clone

how-to

Replace swaps the face. Move takes your head, arms, weight shifts and micro-expressions and gives them to the clone.

Seven steps for the viral 1M-view effect: bullet-list storyboard (position, angle, mood; stick figures optional); film base footage and multi-angle headshots on an iPhone 16 Pro Max; build the base clone in Nano Banana Pro on fal.ai; generate every scene still via ChatGPT prompts plus face swaps; animate scene to scene in Kling 2.6 with motion prompts; motion-transfer the real footage onto the clone with Wan 2.2 Animate in MOVE mode (explicitly not Replace, which 'only swaps the faces'); generate a Kling transition bridging the transferred clip into the animated sequence; edit in CapCut/Premiere with zoom-into-action cuts that hide face seams, a compound clip ('nest') for one global speed ramp, sound design from the Content Creator Templates plugin, Topaz upscale, 4K export.

Do it in this order
Why it matters

One real clip plus AI equals the effect clients see going viral - and the exact mode setting that makes it work.

06

Sam's animated music video: animatic in, master prompt, Hailuo preset, salvage the good two seconds

The animation studio quoted more than the song would ever earn. ChatGPT read his old sketches, Nano Banana drew the cast, Hailuo moved them.

Stack: ChatGPT for prompts, Nano Banana Pro for images, Hailuo (Minimax) for animation, Canva for manual tweaks, CapCut to edit; no upscaling - native 4K/1080p. He fed the years-old animatic MP4 to ChatGPT to extract characters and story, correcting misreads through Q&A (an 'angry dog' read as wholesome). One locked Pixar-style master prompt reused and simplified per character; front-view-only character sheets; environments built frame by frame from the sketches with image-to-image blending. Workflow: image gen, video gen and CapCut in parallel. Hailuo preset: 'no exposure change, make video smooth...'; he avoids Hailuo's end-frame-to-start stitching (visible exposure jumps). Failures: a 'walking past flowers' shot saved by mining a usable 2-second sub-clip; a flower pluck and a shooting star that took many attempts. CapCut beat markers exist, but he cut by feel.

Why it matters

An honest account of the iteration cost - the part every polished demo hides.

07

The edit is the human differentiator: Resolve, CapCut, Premiere habits from all four builds

Stitching is where 'low-effort AI content' turns into something that 'feels more like cinema'.

DaVinci Resolve: free, one-time pro fee versus CapCut/Adobe subscriptions; trim pauses; waveform sync by nudging the playhead; split and delete generated audio so only the ElevenLabs track survives; push video tracks down so baked dialogue never covers narration; crossfades. CapCut/Premiere: zoom-into-action keyframes timed to mask face seams; compound clip / nest for one global speed ramp (fast through transitions, slow on action beats); the Content Creator Templates plugin for a deeper SFX library in Premiere; CapCut beat markers for music cuts. The 14-Day Filmmaker program and the course's editing section are the deep dives.

Why it matters

These are the specific timeline moves that make AI seams disappear.

Tools referenced

ToolCoverageMomentContext
ChatGPTdemonstratedScripts, prompt templates, animatic extraction
PromptEditdemonstratedElevenLabs, Suno, Seedance, Kling front end
ElevenLabsdemonstratedV3 director cues; Voice Changer; SFX tab
SunodemonstratedVocal-free suspense track
Seedancedemonstrated2.0 multi-shot timestamped generation
Klingdemonstrated2.6 product/reaction/minors; motion prompts; transitions
Veodemonstrated3.1 dialogue; first/last frame; under-18 refusal
Nano Banana ProdemonstratedCharacters, edit chains, face swaps
HiggsfielddemonstratedHosts Nano Banana and the Veo/Kling picker
fal.aidemonstratedNano Banana, Kling 2.6, Wan 2.2, Topaz
Wandemonstrated2.2 Animate Move mode
HailuodemonstratedMinimax animation with a stability preset
DaVinci ResolvedemonstratedPrimary editor for both tutorials
CapCutdemonstratedClone and music-video edits
Adobe Premiere ProdemonstratedSFX via plugin; nest
Google DocsdemonstratedStoryboard; Web Page export
TopazmentionedUpscale before 4K export
CanvamentionedManual image tweaks; storyboard sketches

Session materials

Archived locally on V: — click to open. Companion pages link to the LMS.

Action items

    Resources mentioned

    Resources
    • docFree downloads referenced
    • docNot in the caption archive

    Extraction notes

    This page was built from an auto-generated transcript, which garbles product and people's names. Those were corrected silently in everything above and logged here for transparency. The warnings flag claims that were true on the recording day but change fast.

    Transcript corrections applied

    The transcript saysThe trainer actually means
    Sea DanceSeedance
    HigsfieldHiggsfield
    File AI / fall AI / Follow AIfal.ai
    Halo / Hilu / Hylo MinimaxHailuo (Minimax)
    Dr. SquashDr. Squatch
    Mount VeluviusMount Vesuvius
    window mutantswindow mullions
    I'm Lauriepossibly a second presenter; unresolved

    True on recording day — verify before relying