PTENES
TRACK 5

🤖 Visual AI and Generation

The course’s longest learning path: here you master the generation toolkit—Freepik, Seedance, Kling, Luma, and Runway—and learn to control movement, physics, identity, and atmosphere. From YAML structure to particle systems, you direct AI the way a filmmaker directs a crew.

Direction the filmmaker Freepik Seedance Kling Luma Runway The scene executed
6
Modules
~40
Topics
~4.5h
Duration
Advanced
Level

Trail map

Detailed content

5.1~45 min

🧰 AI Tool Pipeline

Cinematic quality doesn't come from AI — it comes from how you control visual language. Learn the role of Freepik, Seedance, Kling, Luma, and Runway, and the three controls that support everything: attention, composition, and connection.

What it is:

How you control the viewer’s attention, build the composition, connect shots, and use the tools correctly.

Why learn:

AI executes—these four controls determine whether the result looks like cinema or noise.

Key concepts:

Cinematic style = clarity + control.

What it is:

Every frame has a subject that’s understood instantly; empty space creates meaning, not mistakes.

Why learn:

The viewer shouldn’t have to search — they should understand the scene immediately.

Key concepts:

More space → freedom/scale · Less space → tension/pressure.

What it is:

Freepik = central workspace, model testing, and asset management. Seedance = strong continuity and character consistency.

Why learn:

They’re the foundation of the pipeline: where you build the scene and where the multi-plane narrative unfolds.

Key concepts:

Node-based pipeline · multiplane storytelling.

What it is:

Kling = dialogue and identity; Luma = realistic movement and atmosphere; Runway = precise control and VFX workflows.

Why learn:

Using the right tool for each task is part of directing.

Key concepts:

There’s no "best" — there’s the right one for each stage.

What it is:

Clear subject within 1s? Intentional composition? Space used? Consistent palette? Natural movement? Do the shots flow? Right tool?

Why learn:

A verification system prevents the “flat AI result.”

Key concepts:

Frame → attention · Composition → meaning · Editing → flow · AI → support.

What it is:

Cinematic quality is a system, not a model. AI supports the process; direction defines it.

Why learn:

Puts responsibility where it belongs: in your visual language.

Key concepts:

Frame → attention · Composition → meaning · Editing → flow.

View Full
5.2~45 min

⚡ Motion Generation in Seedance

Seedance understands physics, not adjectives. Learn to build movement with Force + Time + Weight, sequence large→small→particles, control the camera, and show emotion through the body — never name it.

What it is:

Every movement is built on force (what drives it), timing (how long it lasts), and weight (how the object behaves).

Why learn:

Seedance interprets physics — describing physics means speaking its language.

Key concepts:

Movement = a controlled sequence of events.

What it is:

Primary action first, secondary movement next, residual effects (dust, smoke, debris) last.

Why learn:

This order is reality’s physical signature — without it, everything feels artificial.

Key concepts:

Debris → dust → smoke · Splash → droplets → mist.

What it is:

Position (low/wide/close-up), movement (pan, tilt, dolly, orbit, shake), and behavior (stable vs. handheld).

Why learn:

The camera is what the viewer experiences—always define its behavior.

Key concepts:

Helmet cam · drone follow · constant pull-back.

What it is:

Emotion appears physically: breathing, micro-movements, eye focus, body tension.

Why learn:

"He's scared" generates nothing; "fingers slipping, eyes widening" does.

Key concepts:

Never state the emotion — show it through movement.

What it is:

Precise durations and partial slow motion create realism; vague descriptions destroy it.

Why learn:

Timing is what makes physics feel like physics.

Key concepts:

Use timing, distance, and angle — observable actions.

What it is:

Everything continues after the main action: dust lingers, smoke leaves a trail, water hangs suspended.

Why learn:

These layers create depth and realism.

Key concepts:

You don’t need to describe everything—just the force, sequence, camera, and reaction.

View Full
5.3~45 min

🧾 YAML Control in Seedance 2

AI doesn’t think—it interprets structure. That’s why we use YAML: camera, action, lighting, and transition blocks, with identity locks and time blocks. Clear structure = predictable results.

What it is:

Instead of one long prompt, we divide everything into blocks: camera, action, lighting, transitions.

Why learn:

AI interprets structure—clear structure = predictable results.

Key concepts:

YAML is a control system, not just formatting.

What it is:

base: matching reference @image1 with strict identity lock locks the face, proportions, and design.

Why learn:

Without it, the model redesigns the character in every shot.

Key concepts:

No deviations, no redesigns.

What it is:

Each block controls one dimension of the scene, isolating its function to avoid conflict.

Why learn:

Separate blocks make it clear what to adjust when something breaks.

Key concepts:

Camera ≠ action · always include lighting.

What it is:

One function per field, simple camera, physical movement (not ideas), always include lighting, time blocks 00_02, 02_04.

Why learn:

If the structure breaks, the result breaks.

Key concepts:

Write the transitions explicitly.

What it is:

Cut → structure; whip pan → speed; smash cut → impact; match on action → fluidity.

Why learn:

Transitions define how the scene feels.

Key concepts:

Control rhythm, emotion, and attention.

What it is:

A good sequence flows: scale → speed → emotion → action → release.

Why learn:

AI approximates—transitions break, and iteration is part of the process.

Key concepts:

The goal is control, not just generation.

View Full
5.4~45 min

🎆 Practical Visual Effects

Six techniques that deliver real realism while avoiding "too much CGI": LED volume, practical fire, cables, rain, miniatures, and physical-digital hybrid. It all comes together in the blueprint Light → Texture → Movement → Atmosphere → Emotion.

What it is:

LED walls project environments behind the actors, creating authentic shadows, reflections, and rim lighting.

Why learn:

The real light perfectly “anchors” the actor in the digital world.

Key concepts:

Precise reflections · authentic rim lighting.

What it is:

Real fire brings chaotic movement and organic sparks; physical water wets textures and creates authentic splashes.

Why learn:

Perfect movement looks fake. Chaos looks real.

Key concepts:

Water physics quickly reveals bad CGI.

What it is:

Cables capture real inertia and imbalance; miniatures bring real dust and tactile textures.

Why learn:

Physical weight is one of the hardest things to simulate in post.

Key concepts:

Cinematic framing makes the brain see a miniature as full scale.

What it is:

Real foreground elements combined with digital backgrounds create depth and parallax.

Why learn:

A real detail in the foreground lends credibility to the digital setting in the background.

Key concepts:

Physical grounding for the entire scene.

What it is:

One function per field, physical focus, defined lighting, and controlled camera — watch for drift, timing, and plastic-looking textures.

Why learn:

AI approximates reality—it doesn’t truly understand physics.

Key concepts:

Iteration is a normal part of the process.

What it is:

Cinematic realism follows a flow: Light → Texture → Movement → Atmosphere → Emotion.

Why learn:

It’s the sequence that organizes any effects scene.

Key concepts:

Uncompromising physical realism.

View Full
5.5~45 min

🐢 Slow Motion & Particle Systems

Overcranking (120/240fps → 24fps) and the fast→slow→fast rhythm that makes time matter. And particles as a secret tool: atmosphere that hides AI morphing artifacts.

What it is:

Recording at a high frame rate and playing back at 24fps stretches time and makes movement more emotional.

Why learn:

Always write 120fps → 24fps / slow motion 5x to simulate realistic overcranking.

Key concepts:

Climax, impact, victory — the audience feels that it matters.

What it is:

Emotional impact comes from contrast — interrupting speed with a dramatic dilation of time.

Why learn:

If everything is slow, nothing feels cinematic.

Key concepts:

Change shots every ~2s to create rhythm.

What it is:

Describe blinking, breathing, eye movement, and tension; keep the same light direction, color temperature, and shadows.

Why learn:

Without lighting continuity, shots feel disconnected.

Key concepts:

Living, not artificial, characters.

What it is:

Smoke, dust, sparks, and haze turn a sterile screen into a living frame.

Why learn:

The real world is never perfectly still.

Key concepts:

Atmosphere · movement · realism · depth · energy.

What it is:

Transitional occlusion, volumetric integration, temporal smoothing, and kinetic layers hide morphing and flicker.

Why learn:

Instead of seeking a perfect transformation, use physics to fool the eye.

Key concepts:

Hide topology · soften edges · fragment silhouettes · redirect attention.

What it is:

The transition works not because of anatomy, but because dissolving shadow, smoke, and fur mask the proportions.

Why learn:

Particles are a functional tool that replaces expensive CGI with atmosphere.

Key concepts:

Describe volumetric light, density, turbulence, and interaction.

View Full
5.6~45 min

🎬 Integration with Luma and Runway

With Luma’s Uni-1, you don’t write commands — you converse, describe, and direct. A Unified Intelligence environment that remembers the project context: character sheet → direct a shot → animate.

What it is:

A creative filmmaking environment that keeps characters, lighting, style, and tone consistent across shots.

Why learn:

Stop switching between tools — the project stays cohesive.

Key concepts:

You act as a director, not a prompt engineer.

What it is:

You don’t write commands—you describe and direct naturally.

Why learn:

Speak naturally, and the AI understands your intent, remembers the context, and lets you stay in creative mode.

Key concepts:

"Keep everything the same, just add a cowboy on the left."

What it is:

Ask for a full-body sheet with front, side, and back views—same face, same clothes.

Why learn:

AI can start with just one view—guide it step by step.

Key concepts:

Iteration is essential.

What it is:

Think like a director: a drone shot descends to a cowboy riding between cliffs.

Why learn:

You control camera movement, composition, lighting, and emotion.

Key concepts:

Refine naturally between generations.

What it is:

"Camera like a fast drone circling the rider, high energy, realistic movement" — and refine the jump.

Why learn:

You can compare Kling vs Ray 3.14 and choose what fits your vision.

Key concepts:

"Improve the cinematic flow."

What it is:

When you speak naturally, AI understands your intent, remembers the context, and you stay creative.

Why learn:

You’re not creating prompts—you’re directing your creative partner.

Key concepts:

Sheet → shot → animation.

View Full
← Track 4: Cinematography Track 6: Sound, Editing, and Post-Production →