AI Filmmaking Program · Course 6
In the nine lessons that wrap up the program, you’ll bring together everything you’ve built so far: give each scene a purpose, create sequences that move people, keep your character recognizable from beginning to end, produce a scene with several matching angles, and leave with the plan and production for your first truly finished project.
Lessons
Find out why nine out of ten beautiful scenes still feel empty—and the four-part formula that fixes it.
Six beats that turn disconnected shots into a progression the audience can feel, from the beginning through the scare to the relief.
Turn the six beats into linked generation prompts—without the character’s face changing halfway through the scene.
The preproduction package professionals put together before generating anything—and why skipping this step costs you later.
The reference system that keeps your character from “melting” between one scene and the next.
How to create full scene coverage—wide, medium, close—without making the angles look like they were shot on different days.
The one-page plan that separates those who finish the project from those who disappear halfway through.
The rule that helps you decide on the spot whether to accept a generation or try again—without losing the whole afternoon on a single shot.
Bring together the shot, character, coverage, and production into an assembled film — and plan your final project.
Course 6 · Lesson 1
By the end of this lesson, you’ll test any scene in your script with four questions and know for sure whether it deserves to be in the film.
You’ve seen a technically perfect video—good lighting, steady camera—that still couldn’t hold anyone’s attention. The problem is almost never the image: it’s that the scene isn’t doing any work in the story. This lesson fixes that before you spend a single generation.
watch this lesson on video (English · optional)
↓ role to study
Many people who are just getting started with AI filmmaking believe that “cinematic” means a beautiful image: good lighting, slow motion, film-like color. But a scene can get all of that right and still leave viewers cold. The difference between the two has a name: scene purpose. Before writing or generating any scene, the professional asks this first: why does this scene exist?
The social media manager feels this firsthand: the client asks for "a video of the product on the table," she delivers an impeccable image — soft light, the right angle — and the video holds no one's attention for more than two seconds. It’s not lacking technical quality. The scene just needs to be doing some work: moving the story forward, building emotion, raising tension, or revealing who the character is.
Here’s the most honest test there is: if you removed this scene from the whole video, would anything change? If the answer is no, the scene has no purpose — it just fills screen time.
A scene without a purpose is decoration. A scene with a purpose is a story.
Every scene that works follows the same progression, in any film, on any budget: purpose (why it exists) → objective (what the character wants, right now) → conflict (what makes it difficult) → emotional change (what changes in feeling from beginning to end). This is the backbone behind the old three-act structure from cinema — only shrunk to fit inside a single scene.
Without an objective, the scene is passive: a character just existing, with no direction at all. Without conflict, nothing makes the path difficult, and the audience doesn’t care—the harder success seems, the more the audience roots for it. And without an emotional change, the scene ends exactly as it began, which is the definition of wasted screen time.
Apply this to a small, real script: a social media manager plans a 15-second video about a coffee shop reopening after a month of renovations. Purpose: show the owner’s relief. Objective: open the doors before the morning line loses patience. Conflict: the espresso machine won’t turn on. Emotional change: from panic to pride — in 15 seconds.
Without a concrete objective here and now, the scene becomes a character who is only thinking or simply existing—and the audience feels that it’s a video with no direction. The objective gives it direction; the conflict gives the audience a reason to root for the character. The conflict can be external (tight budget, tight deadline, competition) or internal (doubt, shame, fear of making a mistake in public)—and the harder success seems, the more the audience cares.
A social media manager sees this contrast every week. She films a clothing store owner simply “arranging the window display”—and no one stops to watch. Change the goal and the conflict, and the same scene becomes something else.
Before
The shop owner arranges the display window calmly, without rushing. Nothing pushes the scene anywhere — it’s just a task being done.
After
The shop owner rushes to arrange the display window against the clock: the new collection needs to be ready before the store opens, and the delivery truck is half an hour late.
The payoff: the same action, zero extra production cost—the entire difference comes from giving the scene a goal with a deadline and a real obstacle.
The final piece is the emotional shift: the scene needs to end somewhere different from where it began. Confidence turning into doubt, fear turning into courage, loneliness turning into connection—the label matters less than the movement itself. A scene that feels emotionally identical from the first second to the last is, in practice, a scene that stands still.
Before approving any scene, five questions clear up any doubt: why does it exist? What does the character want? What makes success difficult? What changed emotionally? Why should the audience care? If removing the scene changes nothing in the story, it probably isn’t needed.
And there’s one final test, the hardest of all: mentally remove the dialogue from the scene. Can you understand the emotional change through the action alone — without a single line? If the answer is yes, the scene’s purpose is solid, even without a word.
Test yourself
A scene shows a character cooking calmly from beginning to end, with nothing at stake. According to this lesson’s formula, what’s missing first?
Practice now 0/3 done
Leave with three scenes for your next video, each with a defined purpose, objective, conflict, and emotional shift—in ~10 minutes.
This is planning on paper or in your phone’s notepad—nothing is recorded or generated yet. A mistake here costs one crossed-out line, not an afternoon of lost work.
You have three scenes with defined purposes, without spending a single generation—the foundation the next lesson’s sequence will use.
Summary
Course 6 · Lesson 2
By the end of this lesson, you’ll write a six-beat sequence for one of your short scenes—in the exact order that builds tension, instead of one beautiful shot after another.
AI-generated clips already look beautiful these days—that’s no longer the problem. What’s still missing is the stitching between them: each shot looks like it was cut from a different video. This lesson teaches you the order that turns separate shots into a sequence the audience experiences as a whole.
watch this lesson on video (English · optional)
↓ role to study
Many people who start generating video with AI think cinematic means “a shot that looks like a movie”: these days, almost any generator can deliver that. The real problem is something else: the shots feel disconnected. Each one may look beautiful on its own, but there’s no thread tying one to the next—they look like they were cut from different videos.
Before generating the next shot, people who do this rarely ask, “What’s a cool shot to do next?” The right question is different: what emotion should the audience feel next? A real sequence isn’t a collection of beautiful images — it’s a progression where tension, camera, and rhythm evolve together, from the first second to the last.
The social media manager recognizes the symptom: she delivers three beautiful clips to the client, generated separately, and when they put the three together in the editor, the video looks like three different ads pasted together — not one cohesive piece.
A beautiful shot is a still image. A sequence is an image feeling something happen.
A sequence always follows the same six-beat formula: Establish → Connect → Disrupt → React → Reveal → Escape. It works because tension builds the way the audience naturally processes a scene—first they understand the space, then they get attached to someone, and only then do they feel the danger.
The camera follows this evolution: shot open establishes scale, shot medium builds connection, the close builds tension, and the dynamic movement creates urgency. Changing the order — for example, starting in close-up without ever showing the space — leaves the audience without the grounding they need to care.
Picking up the café video from the last lesson: Establish (the empty store early in the morning) → Connect (the owner’s hands preparing the machine) → Distort (the machine sputters) → React (her face, tension) → Reveal (the steam finally rises properly) → Escape (the door opens, the first customer smiles).
Test yourself
In the sequence of six beats, what comes right after “Establish”?
There’s an order-of-operations rule that matters more than it seems: always show the character's reaction before the problem. When the audience sees the worried face first and only then understands why, the problem feels heavier—because we already sense that something matters before we know what.
Working backward is the most common mistake: showing the threat or problem right away, with no setup, strips away the impact it should have. The audience sees what happens before they care about who it happens to.
Before
Cut straight to the broken espresso machine. Only then show the café owner’s reaction.
After
Cut first to the café owner’s tense face—she hears a strange noise. Only then show the machine smoking instead of steaming.
The payoff: the same two shots, in reverse order—the version that shows the reaction first makes the problem feel real.
Four mistakes show up again and again, in this order of frequency: jumping straight to close-ups without establishing the space; revealing important information too soon; speeding up the pacing too quickly without letting the tension build gradually; and ending the sequence abruptly right after the strongest moment, without giving the audience a second to exit.
A neighborhood pizzeria owner planning the shop’s anniversary video fell into this last mistake: they cut straight from the dish coming out of the oven to the pizzeria’s logo — without showing anyone tasting it, without a second of reaction. The scene had all the right ingredients and ended up saying nothing, because it was missing “Escape” — the exit moment that completes the feeling.
And here’s an honest caveat: in very short clips, 5 to 8 seconds long, there isn’t always room for all six beats. It’s fine to skip beats — what you can’t do is change the order of the ones you keep.
In the end, every good sequence is a line of feeling rising and falling with intention: curiosity → connection → tension → relief (or surprise, or pride — it depends on the story). The job of the person sequencing is not to choose beautiful shots; it's to decide where that line rises and where it falls.
The six-beat formula works for any profession telling any small story—the only thing that changes is the context.
Apply it to your profession
Practice now 0/4 done
Leave with a beat-by-beat script for a 10–15-second sequence, ready to guide your next generations—in ~10 minutes.
This is still just a written script—nothing is generated in this exercise. Changing the order of the beats means rewriting one line, not reshooting anything.
You have a complete sequence of six beats written — the script that the next lesson turns into a series of linked generation prompts.
Summary
Course 6 · Lesson 3
By the end of this lesson, you’ll generate two consecutive beats of your sequence with the same character, lighting, and setting—without the face changing from one generation to the next.
You already wrote the six-beat sequence in the previous lesson. The risk now is technical: every AI generation starts from scratch, and without care, the character’s appearance changes by the second beat. This lesson gives you the method for keeping everything consistent, generation after generation.
watch this lesson on video (English · optional)
↓ role to study
You already know from earlier courses in the program that every AI generation starts from scratch: it doesn’t remember the face it created in the previous generation. That applies just as much to a sequence of six beats. The temptation is to write one huge prompt describing all six at once—and that’s exactly where the character’s face changes along the way.
The professional approach is the opposite: each beat becomes a separate generation, and they all attach the same reference image—the same photo of the character, setting, and lighting you’ve been using since the grid method back in course 1. It’s the reference, not the text, that carries an identity from one generation to the next.
For the social media manager, this changes the whole routine: instead of hoping all six generations happen to match, they generate a beat, check it, attach the same reference again, and generate the next—with the control that comes from knowing exactly what they’ll get.
Video generators respond better to direct camera directions than to vague descriptions of “mood.” For each beat, describe the movement along with the shot: “slow, wide pan” for Establish, “medium push-in” for Connect, “locked-off close-up” for React, “quick, abrupt movement” for Escape. The beat’s urgency determines the movement—not the other way around.
Before you see the finished practice prompt, think of it as an order form with four fixed fields—beat, shot, action, and camera. You fill in each field with a short sentence; the rest of the text around it stays fixed, and you don’t need to rewrite it.
Common mistake
Write all six beats in a single prompt, expecting the AI to assemble the whole sequence on its own. The generator treats this as one confusing scene and mixes elements from all six beats into an incoherent image. Each beat needs its own generation, one at a time.
Most generators have a control for how much the AI can “make up” on its own—the higher it is, the farther it strays from what you asked for; the lower it is, the more closely it sticks to the script. Think of it as the reins on a guided horse: short reins, and it follows the trail exactly; loose reins, and it wanders through the field however it thinks looks nice.
For a connected sequence, keep a tight leash: it keeps the generator close to your description and reference image instead of reinventing the light or framing with every beat. A loose leash is for exploring individual ideas—not for tying together a six-part sequence.
A social media manager tested this while putting together the lost box sequence: with loose reins, each beat came out with a different colored light; with tight reins, all six beats looked like they were filmed that same morning, in the same store.
Putting it all together: a short, repeatable workflow turns the written sequence from lesson 2 into a real generated sequence, without the character changing along the way.
The workflow, beat by beat
Test yourself
When generating beat 3 of your sequence, which image should you attach as a reference?
Practice now 0/4 done
Leave with two generated frames—from consecutive beats in your sequence—with the same character, lighting, and setting in both—in ~12 minutes.
Each generation is independent and doesn’t ruin the previous one: if a frame comes out wrong, generate it again without losing anything that worked.
Using the attached reference image to lock character, wardrobe, environment and lighting exactly as shown, generate the next shot of this sequence: BEAT: <qual das seis batidas: Establish / Connect / Disrupt / React / Reveal / Escape> SHOT: <wide shot / medium shot / close-up> ACTION: <o que acontece nesta batida, em uma frase> CAMERA: <static / slow pan / tracking / whip pan — conforme a urgência> Keep the exact same face, clothing, environment and lighting as the reference image. Do not redesign anything. No new elements, no stylization changes.
You have two frames from your sequence, generated and connected by the same character and lighting—the beginning of your first sequence-guided stretch of film.
Summary
Course 6 · Lesson 4
By the end of this lesson, you’ll put together a complete preproduction package—idea, references, character and setting descriptions, and a key frame—before spending a single generation on your project.
Whoever generates without any preparation pays for it later: a character whose face changes, a setting whose color changes, an entire afternoon spent generating variations with no direction. What professionals do differently isn’t write a better prompt—it’s decide all of this beforehand.
watch this lesson on video (English · optional)
↓ role to study
People who are just starting out tend to open the tool, write a request, and hope the AI will handle everything on its own. The result of practicing this way is always similar: inconsistent characters, settings that change from scene to scene, random lighting, and no creative direction. The difference between an amateur-looking AI video and one that looks like a real production is almost never in the final prompt—it’s in the decisions made before it.
Strong video comes from a strong workflow, not a lucky prompt. Before any generation, someone has already decided who the character is, what they feel, and what’s happening. Those decisions, made deliberately, are worth more than any last-minute tweak to the prompt.
A neighborhood pizzeria owner planning a video for the shop’s 10th anniversary learned this firsthand: they opened the generator right away and asked for “a beautiful video of the pizzeria” — then spent the whole afternoon testing variations without any direction, because they had never decided what the video needed to show.
A good prompt comes from decisions made beforehand. Without those decisions, not even the perfect prompt can keep the film on its feet.
Pre-production is the name for this work of deciding in advance: the story, a visual reference board, the character, and the setting — all written down or gathered before the first generation.
The reference board brings together photos, colors, and moods that serve as a compass for the entire project—some tools call this a "moodboard," but the idea is simple: a board of saved images (on your phone, in a folder, or in a notes app) that shows the mood you want to achieve, so you can refer to it whenever you’re unsure.
The pizzeria owner did this before generating anything: collected photos of bakeries with warm lighting, worn wooden tables, and family portraits on the wall. This board became the video's visual compass for the entire anniversary — every generation afterward was compared against it.
The second part of the foundation is to write, just once, what the main character and setting look like: appearance, clothing, colors, weather, lighting. That description becomes the reference every subsequent generation will follow—instead of being reinvented with each new request.
This is where the biggest time savings happen, and the contrast is clear side by side.
Before
The pizzeria owner generates ten variations of the chef with no fixed description, hoping one of them will look right — none match the previous one exactly.
After
He writes it once: "50-year-old male chef, white apron with a flour stain on the chest, graying mustache, tired but proud smile." Every following generation uses that same description.
The payoff: from crossing your fingers through ten generations to get one that works, to a written description that guides every next attempt with confidence.
Before producing the entire scene, generate three keyframes representing the beginning, middle, and end of the story—as a quick proof of concept. They test whether the lighting, composition, and mood match what you imagined before you commit time to the full production.
The pizzeria owner generated three frames: the oven off late at night, the oven lit with the first batch baking, and the dining room packed on the night of the anniversary. Together, the three confirmed that the warm light from the reference board really worked on screen—before he generated the whole scene.
This costs three generations, not thirty. It's much cheaper to find out now that the lighting doesn't fit than to discover it midway through production, with half the project already generated.
Putting it all together: pre-production is the story, plus the reference board, plus the character, plus the setting, plus the key frames—all decided in writing once, before the first generation. Without this preparation, you get random generations and hope they turn out well. With it, you gain real creative control.
The final test before moving on: if you removed pre-production from the process, would you still know, without hesitation, who the character is and what the scene’s lighting should be? If not, the foundation isn’t ready yet.
Common mistake
Skipping pre-production because “I just want to see something finished already.” The 15 minutes this step takes pay off many times over—people who skip it usually spend an entire afternoon regenerating the same character and trying to make them look consistent.
Test yourself
According to this lesson’s formula, what comes BEFORE generating any keyframe?
Practice now 0/4 done
Walk away with a mini preproduction package—an idea, two references, a described character and setting, and a key-frame concept—ready to guide your next production, in ~12 minutes.
Everything here is a written decision — no generations are made in this exercise. The worst-case result is a line you rewrite in 30 seconds.
You have a complete pre-production package—the foundation that keeps you from regenerating everything from scratch when something doesn’t match.
Summary
Course 6 · Lesson 5
By the end of this lesson, you’ll have generated a master character and a second image of them in a different pose, with a face, outfit, and proportions identical to the original.
Generating a beautiful character once is easy—any generator can do that today. The hard part is keeping them the same person in the fifth or tenth generation. Without solving that, your final project doesn’t have a cast; it has a succession of similar-looking strangers.
watch this lesson on video (English · optional)
↓ role to study
Creating an impressive portrait once is easy. The real challenge begins when that same character needs to appear in several different scenes, emotions, and settings: after a few generations, their face, clothes, proportions, and hairstyle start changing without you asking.
One consistent character is exactly the opposite: the same recognizable visual identity, scene after scene, even when none of them were generated together.
The pizzeria owner learned this firsthand: he created a chef mascot for the restaurant’s posts, and the first image came out perfect—but by the fifth post, the mascot’s face had changed enough for a loyal customer to joke in the comments: “Is this our chef’s twin brother?”
The professional solution starts with a master character: a single high-quality, detailed image that becomes the official character reference for the entire project. From it, you can also generate an emotion sheet — the same face in several different expressions, preserving the identity in each one.
The key is how you use this image afterward: instead of describing the character in words with every new prompt, you attach the master character image as a visual reference. Text alone drifts with each generation; an attached image anchors it.
Common mistake
Describing the character in words every time instead of attaching the master character image. Two descriptions in Portuguese never come out exactly identical word for word — small text variations lead to small variations in the face. The attached image solves that once and for all.
Along with the master character, it’s worth creating a reference sheet with the main angles: front, profile, three-quarter, and full body. You’ve seen this idea before in the program—the grid method in course 1 and the angle sheet in course 2—but now it takes on a different role: it becomes the official document of the character, not a study draft.
With all four angles ready, the AI gets complete visual information about the face and body—instead of guessing what someone’s profile looks like when it has only seen them from the front.
The pizzeria owner put this sheet together just once for the chef mascot and saved it as the character’s “official document”—today, it’s reused for every new post, and now it’s the foundation for the anniversary film.
Along with the reference images, a few words in your prompt help lock in the identity: repeating phrases like "exact same face, same features, consistent character" reinforces in text what the image is already showing. It’s reinforcement, not a substitute — the reference image still does the main work.
Combining the pieces from earlier, the complete workflow always follows the same order.
The workflow, in the right order
Once the reference set is ready, it can be used for any scene: different environments, actions, and camera angles—while always keeping the same facial structure, hairstyle, clothes, and body proportions.
In the pizzeria's anniversary film, the chef mascot appears in the kitchen at dawn, in the packed dining room at night, and at the door welcoming the first customers — three very different settings, and still recognizable in all three, because every generation pointed to the same reference set.
Test yourself
You’re going to generate your character in a new scene. What ensures greater consistency: describing everything again in text, or attaching the reference set?
Practice now 0/4 done
Walk away with a master character image and a second generation of that character in another pose, with the same face and clothing—in ~12 minutes.
The master character image is saved separately—no new generation deletes or alters the original. If the second image looks different, the master remains untouched so you can try again.
STEP 1 — Master character: Create a highly detailed, photorealistic character portrait to use as a master reference: <descreva aparência: idade, rosto, roupa, cor marcante>. Sharp focus, consistent studio lighting, plain neutral background, front-facing. STEP 2 — New pose, same identity (attach the master image): Using the attached reference image, keep the exact same face, hairstyle, clothing and body proportions. Generate the same character in a new pose: <nova pose ou ação, em uma frase>. Consistent character, same identity, no redesign.
You have a master character and a second generation that stays true to it—the reference foundation that will support every scene in your final project.
Summary
Course 6 · Lesson 6
By the end of this lesson, you’ll create a list of three shots for one of your scenes—wide, medium, close—with the same character and lighting confirmed in all three.
Even the best single shot leaves the editor with no options when it’s time to cut. Without several angles of the same moment, every scene in your film becomes a standalone postcard — beautiful, but with no possible editing rhythm.
↓ role to study
In film, shooting the same moment from more than one distance and angle has a name: scene coverage. Without coverage, the editor is stuck with a single angle from start to finish—and the whole scene feels flat, with no possible cutting rhythm.
The pizzeria owner planning the toast at the anniversary dinner needs more than a beautiful image: he needs the wide shot (the packed dining room), the medium shot (the celebrating table), and the close-up (hands raising glasses)—three angles of the same moment, not three different moments.
Without all three, the edit has only one cut option—and a film with one cut option per scene feels like a presentation slide, not cinema.
Before generating the first shot of the scene, write the full list: how many shots, from what distance, showing what. This short list avoids the most costly mistake at this stage—generating one shot at a time on impulse, without checking that it uses the same assets as the previous ones.
Common mistake
Generate every shot in the scene from scratch, without attaching the same master character and lighting description used for the other shots. The result is three beautiful images that look like they were taken on different days—the audience feels the break even if they can’t name why.
The character isn’t the only thing that needs continuity — the light and the environment do too. In course 2, you saw how to relight an image without redoing the scene and how a production sheet organizes each decision by field; here, those two tools are back, now applied to the multiple shots in a single scene.
Describe the light the same way in every shot: if the wide shot has “warm candlelight, late at night,” the close-up needs the same phrase—never “neutral studio lighting” just because the close-up came from a separate prompt. The pizzeria owner learned this when he noticed that the close-up of the toast had office lighting, while the other two shots had the dining room’s warm light—repeating the same lighting description was all it took to make the three shots match.
Combining the shot list, the character assets from lesson 5, and lighting continuity, the complete workflow for a coverage scene always follows the same order.
The workflow, in the right order
One honest point: the rule isn’t “always generate all three complete shots” — it’s “cover what the scene needs to feel like one continuous moment.” Very short scenes sometimes work better with just two shots, chosen for the contrast they create, not their number.
For the toast shot, the pizzeria owner decided to use only the wide shot and close-up, skipping the medium shot—the contrast between “the entire dining room” and “hands raising a glass” carried the emotion on its own, with no middle step needed.
Test yourself
You generated three shots of a scene, but the close-up came out with different lighting from the other two. What’s the most likely cause?
Practice now 0/4 done
Walk away with a list of 3 shots (wide, medium, close-up) for a scene in your project, with the character and lighting assets confirmed—in ~12 minutes.
This is still just a written list—nothing is generated in this exercise. Reordering the shots means reordering a list, not reshooting anything.
You have the shot list for an entire scene, with continuity assets already confirmed—ready for actual generation.
Summary
Course 6 · Lesson 7
By the end of this lesson, you’ll fill out a one-page plan for your final project — with the scene, sequence, coverage, and character confirmed — ready to guide your production without losing any decisions along the way.
You already know how to decide on purpose, sequence, consistent characters, and coverage—separately, lesson by lesson. What usually derails a project isn’t a lack of skill: it’s having those decisions scattered in your head, in phone photos, and in random messages until they disappear right in the middle of production. This lesson brings everything together on one page.
↓ role to study
In the previous six lessons, you learned to decide on purpose, sequence, consistent characters, and scene coverage—each piece in its own lesson, each one worked out on its own. The problem comes when it’s time to bring everything together: if each decision lives somewhere different—a photo saved here, a voice memo there, and the rest only in your memory—it gets lost right in the middle of production, when you need it most.
O one-page plan solves exactly that: all the decisions for your final scene, gathered in one place, ready to consult without having to reconstruct anything from memory.
The photographer felt this gap firsthand: she finished a 20-second film for a jewelry brand—the moment the customer opens the little box and sees the piece for the first time—but the character reference was in a folder, the coverage list in a note, and the sequence only in her head. On generation day, she spent twenty minutes just looking for the right reference.
A plan in your head is a promise. A plan on a page is a commitment anyone can follow — including you, two weeks later.
The one-page plan always has the same four sections, each prompting a decision you already know how to make: scene and purpose (what changes emotionally and why, from lesson 1) → condensed sequence (the order of the beats, even reduced to a few lines, from lessons 2 and 3) → coverage list (how many shots, from what distance, from lesson 6) → confirmed assets (which master character, which lighting and environment reference, and where they’re saved, from lessons 4 and 5).
The photographer filled out the jewelry film’s shot plan like this: scene—open the box and reveal the piece, from anticipation to delight; sequence—establish the table with the box closed, connect with hands opening it, disrupt with a second of hesitation, reveal the shining jewelry, escape in the smile; coverage—a wide shot of the table, a medium shot of the hands, a close-up of the jewelry and the reaction; assets—the master character reference for the model’s hands, and the warm afternoon shop light repeated across all three shots.
The difference between scattered decisions and consolidated decisions isn’t subtle — it shows up on the clock, at the exact moment you need speed most: production day.
Before
The final scene’s decisions live in loose photos in the phone’s camera roll, a voice message, and what the photographer remembers. When it’s time to generate, she has to piece everything together in a hurry.
After
The same four decisions, written down once in a single note — scene and purpose, sequence, coverage, assets. When it’s time to generate, she opens the note and knows exactly what to do.
The payoff: from spending twenty minutes looking for the right reference to spending ten seconds opening a note—the same decision, a different format.
There’s a simple test to see whether your one-page plan is complete: if someone else had to produce the scene using only this page, without being able to ask you anything, could they do it? If the answer is no, some section is still too vague—and it’s worth identifying which one before moving on.
The photographer applied this test to the jewelry film’s shot and realized the assets section only said “warm light”—without saying what time of day it was or whether it came from a window or a lamp. She rewrote it as “warm afternoon light coming in from the side through a window”—specific enough for any generation to follow without uncertainty.
Test yourself
What’s the right test to tell whether your one-page plan is ready?
Practice now 0/4 done
Leave with all four sections of the plan filled in for your project’s final scene—scene and purpose, summarized sequence, coverage, confirmed assets—in ~10 minutes.
You just organize decisions you've already made in earlier lessons — nothing new is generated here. If you change your mind later, just edit the same note at no cost.
You have the one-page plan for your final project ready—the document that guides your production without letting you lose any decisions along the way.
Summary
Course 6 · Lesson 8
By the end of this lesson, you’ll apply the three-attempt rule to a generated shot from your project and decide, without getting stuck, whether to accept it, adjust it, or change the shot.
Without a clear rule, a generation day becomes a bottomless pit: an entire afternoon stuck on the same shot, trying again without knowing when to stop. This lesson gives you the criterion to decide in the moment, without drama and without losing the rest of your coverage list.
↓ role to study
You already have the one-page plan from the previous lesson. The remaining risk is different: without a stopping rule, generating becomes a loop with no exit—each almost-right result invites you to try one more time, and the whole afternoon disappears into a single shot.
The photographer experienced this on the jewelry film’s production day: she spent forty minutes just on the close-up of the piece being revealed, a dozen attempts, the light changing a little each round, with no criteria for knowing whether to stop or continue. The rest of the shot list got squeezed into the end of the day.
The problem is rarely that the tool generates poor results. It’s the lack of a clear rule for when to stop trying and when to move on.
A final production is where you generate for real, intending to use the result — no longer a test like the keyframes in lesson 4. That’s why it calls for a firmer criterion than “I like it” or “I don’t like it.”
This criterion comes straight from your one-page plan: does the generated shot match the character, lighting, framing, and action you decided on? If all four match, accept it. If one doesn’t, you know exactly which one — it’s not a vague feeling, it’s a specific item on the list that failed.
The photographer started using exactly these four questions on the monitor, shot by shot: “is this the reference hand? is this the agreed afternoon light? is this the medium shot on the list? did the box get opened?”—four yeses, approved; any no, and she knows what to adjust.
Combining the acceptance criteria with a limit on attempts, the final production day workflow always follows the same order:
The workflow, attempt by attempt
That’s the workflow the photographer followed for the jewelry close-up: three tries with small adjustments, the light kept the same, and on the fourth round—instead of insisting—switched to the backup shot already noted on the coverage list.
Here’s a fine distinction: the goal of the final production isn’t to match the keyframe one hundred percent — it’s to serve the scene’s purpose, the one from lesson 1. Small variations no one notices on the final screen are acceptable; the wrong character or an action that doesn’t happen never is.
The photographer accepted the wide shot of the table even though one shadow was slightly different from the plan—it didn’t affect the scene’s purpose, which was to show the anticipation before the box was opened. But she rejected a close-up where the model’s hand didn’t match the reference, even though the frame looked beautiful on its own.
Common mistake
Rewrite the entire prompt from scratch for every attempt instead of adjusting only the field that failed. Each complete rewrite resets what was already working — including the strength of the attached reference — and makes it impossible to tell later what fixed or broke the result. The tight control you learned in lesson 3 stops working if the text changes completely every round.
The photographer learned this through practice: on the first close-up attempt, she rewrote the entire prompt to try to fix the wrong hand—and the lighting, which was already right, changed too. On the second attempt, she changed only the action field, leaving everything else intact: the right hand appeared, and the lighting stayed the same.
Practice now 0/4 done
Walk away with a shot list plan generated and evaluated against acceptance criteria, using no more than three attempts—in ~12 minutes.
Each attempt is a new generation that doesn’t erase previous ones — comparing three versions side by side takes only a glance, never a lost production.
The prompt below is a form for targeted adjustments: fill in only the field you need to change, without rewriting the rest.
Using the attached reference image(s), generate this shot again with ONE targeted change: SHOT (unchanged from before): <qual plano da sua lista: wide / medium / close-up> FIELD TO ADJUST: <apenas um: lighting / framing / action / expression> NEW VALUE: <o ajuste específico, em uma frase> Keep everything else exactly as the previous version: same character, same environment, same lighting unless stated above. Do not redesign the shot from scratch.
You have a plan evaluated using the three-attempt rule—the criterion you’re ready to apply to the rest of your coverage list without losing the whole afternoon to a single shot.
Summary
Course 6 · Lesson 9
By the end of this lesson, you’ll have your entire final project plan assembled — with the scene, coverage, character, and editing order decided — ready to produce and deliver as a finished film.
Many people generate every shot perfectly and never make it to the end: the pieces sit loose in a folder without becoming a film. Most abandoned projects stop right here, one step away from being ready. This lesson completes that last step—and takes you straight to your own final project.
↓ role to study
A final edit it's not a mechanical step of pasting clips in the order that was already written down. Even with the sequence decided in your one-page plan, when you have the actual generated shots in hand, one may call for reordering, cutting, or replacing—the edit is still an authorial decision, not a fitting-together task.
The photographer had decided on the order of the six beats in the jewelry film’s shot plan, but when she saw the shot for the "hesitate" beat—a second of doubt before the hand opens the box—she thought it disrupted the scene’s rhythm. She cut that beat in the final edit, even though she had already generated the shot, because the scene was stronger without it.
Before calling the final edit ready, a short checklist confirms that nothing was missed:
The assembly checklist
The photographer applied this checklist to the jewelry film and found, in item 2, a wide shot with a shadow slightly different from the other two—small enough to let pass, since it didn’t change the scene’s purpose.
Real-world example
Marina, a photographer with her own studio, landed a 20-second film for a jewelry brand: the client opening the box and seeing the piece for the first time. She filled out the one-page shot plan from lesson 7, generated the four coverage shots using the three-attempt rule from lesson 8 — replacing only the close-up with a backup shot after the third attempt failed to match the light. In the final edit, she cut the extra "hesitate" beat, reviewed the shots side by side using this lesson’s checklist, and watched from beginning to end: anticipation turned into delight, exactly as the plan called for. She delivered the finished film to the client that same day.
Today’s practice is a guided project: brings together the one-page plan, the final production rule, and this lesson’s assembly checklist in a single workflow. Only the planning has a stated time — the actual production, generating and assembling each shot, is up to you, at your own pace, off this screen.
That’s exactly the path the photographer followed, from planning to delivery—and it’s the same one waiting for you next, with your own project, in your own scene.
Practice now 0/5 done
Leave with a complete plan for your final project—scene, coverage, acceptance criteria, and editing order decided—in ~15 minutes; production itself (generating and editing each shot) continues at your own pace, outside this page.
Every decision here can be revisited: reviewing the plan, generating again, or reordering the shots won’t erase anything that already turned out well. Go back and adjust as many times as you need, without rushing.
You have the plan for your final project ready, from beginning to end—the film itself is the next step, on your own timeline, with everything already decided.
Summary