
Kick off your Kling AI video creation journey and learn to generate cinematic AI videos through creativity, exploring short films, YouTube content ads, and animated characters with fast-start guidance.
Kling Step-by-Step takes you through Kling from the ground up, showing you what it is, how it works, and how to use its tools systematically. Rather than jumping straight into advanced features, we build your understanding step-by-step so you know not just what Kling can do, but how to use it effectively.
We start with the fundamentals: getting access to Kling, understanding the different plans, and learning text prompting — one of the core skills you’ll need throughout the course.
From there, we move into images. You’ll learn text-to-image generation, how to create consistent character and set sheets, how to use reference images, and the different techniques you can use to build the visual foundations for your projects.
Once you understand the image tools, we move into video. We cover text-to-video, using references to guide video generation, multi-shot video generation, adding audio to your videos, motion control, and avatars — giving you a complete understanding of Kling’s video capabilities.
Finally, we look at sound effects in Kling, bringing your generated videos further towards a finished piece.
By the end of the course, you’ll have worked through Kling systematically, from the fundamentals through to its more advanced tools, and you’ll understand how to use Kling fully to create images, characters, scenes, videos, movement, avatars, and sound.
Using Kling Natively means working directly inside Kling rather than accessing Kling’s models through another AI tool or platform. While Kling is now available inside a number of other creative tools, those integrations may not expose every feature or control that Kling itself offers.
To make sure you have access to the full range of Kling’s tools and features, throughout this course we’ll be using Kling natively, directly through the Kling platform.
This gives us a consistent starting point and ensures that as we move through the course, you’ll be learning the tools exactly as they’re designed to be used inside Kling.
Plans & Subscriptions takes a closer look at the different Kling plans and subscription options, so you can understand what you’re actually getting for your money.
We’ll go through the available plans, look at how credits work, and break down the costs so you can compare the different options more easily. We’ll also work out the cost per credit and the approximate cost per video, giving you a practical way to understand what your Kling generations are really costing.
By the end, you’ll have a clear understanding of how Kling’s pricing works and which costs to consider when planning your workflow and projects.
The Kling Interface lecture gives you a guided tour of Kling’s layout and workspace, so you always know where to find the tools and settings you need.
We’ll go through the interface step-by-step, looking at the different sections, menus, controls, and workspaces you’ll be using throughout the course.
The goal is simple: to make sure you never feel lost inside Kling. By the end of this lecture, you’ll know your way around the platform and be ready to start using its creative tools with confidence.
This Text Prompting in Kling lecture teaches you how to write effective prompts specifically for Kling, using Kling’s own official documentation as our foundation.
We’ll break down how Kling interprets prompts, what information to include, and how to structure your descriptions to get more consistent and controllable results.
We’ll also use an LLM to help generate and refine prompts, creating a repeatable prompting workflow that is specifically suited to Kling rather than relying on generic AI prompting techniques.
By the end of this lecture, you’ll have a practical prompting system you can use to consistently create Kling-ready prompts for your projects.
Text-to-Image Generation shows you how to turn your ideas into images directly inside Kling. Building on the prompting fundamentals from the previous lecture, we’ll use text prompts to create and control the visuals we need for our projects.
We’ll work through the text-to-image tools ("generate") step-by-step, looking at how to describe characters, environments, compositions, styles, and details to achieve more predictable results.
We’ll also explore how to iterate on generations, refine prompts, and make deliberate changes rather than simply generating images at random.
By the end of this lecture, you’ll understand how to use Kling’s text-to-image generation to create the visual assets you need and have a solid foundation for the image-based workflows that follow.
Character & Scene Sheets puts the prompting techniques we’ve learned into practice, creating the visual references we’ll need for the rest of the course.
We’ll work through a real-world example, building character sheets and scene sheets that establish how our characters and environments should look from different angles, compositions, and perspectives.
These sheets become an important part of our workflow later when we move into video generation. By creating consistent visual references upfront, we can use them to help Kling maintain the same characters, locations, and overall visual identity across multiple shots.
Using Image References shows you how to use the character and scene sheets we created in the previous lecture to build consistent images inside Kling (which we'll turn to video later).
We’ll take our references and combine characters and environments together, learning how to place our consistent characters into different scenes while maintaining their appearance and visual identity.
We’ll also explore how to use references for other scenes, giving us greater control over the composition, setting, style (we can use a style as a reference too), and overall look of our images.
By the end of this lecture, you’ll have a complete understanding of Kling’s image workflows — from creating images with text to using multiple references to build almost anything you can imagine with greater consistency and control.
Text-to-Video Generation takes everything we’ve learned about prompting and images and brings it into motion.
We’ll learn how to use Kling’s text-to-video tools to turn a written description into a moving scene, breaking down how to prompt for action, camera movement, composition, timing, and cinematic detail.
We’ll work through real examples, showing how to move from an idea to a usable video generation and how to refine our prompts when the result isn’t quite what we want.
By the end of this lecture and section, you’ll understand how to create videos from text and have the foundations needed to start building more complex, controlled video sequences.
Image References (elements) to Video shows you how to take the consistent characters, environments, and visual references we created earlier and bring them to life in Kling.
We’ll use our reference images to guide video generation, learning how to maintain the appearance, style, composition, and identity of our characters and scenes while introducing movement and camera direction.
We’ll work through real examples, combining our references with prompts to create controlled video sequences rather than relying on text alone.
By the end of this lecture, you’ll understand how to turn your established visual references into consistent, usable video, giving you much greater control over how your characters and scenes come to life.
Audio in Video explores the different ways we can create spoken dialogue and sync it to our Kling videos.
We’ll compare lip-syncing an existing video with Kling’s ability to natively prompt for speech within a video, looking at how each approach works and when you might use one over the other.
We’ll work through practical examples, testing both methods and seeing how they affect dialogue, lip movement, timing, and overall performance.
By the end of this lecture, you’ll understand the different ways to add speech to your Kling generations and how to choose the right approach for your project.
Start & End Frames gives you more control over how your Kling video begins and ends. You can provide an image for the start frame and another image for the end frame, giving Kling a clear visual target for the shot.
This is especially useful when you need a video to move between two specific moments, compositions, or poses. Instead of leaving the ending entirely up to Kling, you can define exactly where you want the shot to finish.
We’ll test using start and end frame to see how Kling transitions between them, and how we can use prompting to guide the movement in between.
By the end of this lecture, you’ll understand how to use start and end frames to create more controlled and intentional video sequences.
Multi-Shot Video Generation lets you create multiple shots and cuts within a single Kling generation. Instead of generating one continuous shot, you can use a single prompt to describe a sequence of different shots that work together to tell a complete story.
This is especially useful when you want to create a meaningful 15-second scene or even a complete advert without having to generate and edit each shot separately. You can describe the shots, camera changes, actions, and progression you want, and Kling can bring them together into one sequence.
We’ll work through a real example, breaking down a scene into multiple shots and learning how to structure prompts so each cut contributes to the overall story, rather than feeling like a collection of unrelated clips.
By the end of this lecture, you’ll understand how to use multi-shot generation to create complete, purposeful video sequences from a single text prompt — with far less need for traditional editing.
Video-to-Video Editing (transform) lets you transform an existing video while keeping the original motion, composition, and structure as a foundation. Instead of generating a completely new video, you can use Kling to modify what is already there.
This could mean changing a character’s clothing, transforming the environment, adding or removing elements, or completely changing the subjects while preserving the underlying video. This gives you a powerful way to edit and reimagine footage without starting from scratch.
We’ll work through a practical example, making changes and seeing how well Kling can maintain the original movement and format while transforming the parts we want to change.
By the end of this lecture, you’ll understand how to use video-to-video transformation to turn existing footage into something completely different while still retaining control over the original video.
Motion Control in Kling lets you upload a reference video to guide your character’s movement. Instead of relying only on text prompts, Kling analyzes the actions in the reference clip and makes your AI character copy those motions.
This is especially useful when text-to-video or image-to-video prompts aren’t giving you accurate or natural character movement. By using a reference video, you gain much more control over how your character moves and behaves.
Let’s test it.
Avatars lets you create a reusable digital character that can appear across your Kling videos. Instead of recreating a character from scratch each time, you can build an avatar and use it as the consistent subject for different scenes and performances.
This is especially useful when creating recurring characters, presenters, influencers, or branded content where you want the same person or character to appear across multiple videos. You can then combine the avatar with different prompts, environments, movements, and dialogue.
We’ll create and test creating an avatar exploring how they can be used to maintain a consistent character while changing the setting, action, and story around them.
By the end of this lecture, you’ll understand how to create and use Kling avatars to build consistent, reusable characters for your video projects.
Sound Effects in Kling lets you generate audio effects directly from a text prompt to add to your edit and bring your videos to life. Instead of sourcing sound effects separately, you can describe the sounds you want and use Kling to create them for your scenes.
This could be anything from footsteps, doors opening, vehicles, impacts, environmental sounds, or atmospheric details that help make a scene feel more immersive and complete.
We’ll experiment with different prompts learning how to describe sounds clearly.
By the end of this lecture, you’ll understand how to generate sound effects from text in Kling and use them to add another layer of realism and detail to your videos.
Video-to-Sound Effects lets you use an existing video as the basis for generating sound effects. Instead of manually describing every sound, Kling can analyse what is happening in the footage and create audio that matches the actions and environment.
This is useful for adding footsteps, movement, impacts, environmental sounds, and other audio details that naturally fit the visuals. It can help turn a silent video into a much more immersive scene without having to source and place each sound effect yourself.
By the end of this lecture, you’ll understand how to turn your existing video into a more complete audiovisual sequence using automatically generated sound effects.
Walk away confident and inspired to create amazing videos with Kling AI, refining prompts and pushing the boundaries of what AI video creation can do.
What if you could create jaw-dropping videos, complete with lip sync, sound design, and cinematic motion—using just your ideas?
Kling AI is one of the most powerful video generation platforms available today. In this hands-on course, you’ll learn exactly how to use Kling’s advanced tools to turn your words and images into fully animated, professional-quality video content.
Whether you're a content creator, marketer, educator, or simply curious about the future of video production, this crash course will help you master Kling's interface, features, and workflows—no editing background required.
You’ll Learn How To:
Generate high-quality videos using only a text prompt
Create animated visuals using your own images, references, or photos
Add cinematic motion, style, and elements using Kling’s intuitive tools
Use lip sync, AI sound effects, prompt for speech and avatars for polished, pro-level results
Edit, restyle, and refine your generated content
Build short videos for social, marketing, storytelling, or entertainment
By the end of this Kling AI video course, you’ll know exactly how to bring your ideas to life with Kling—from your very first prompt to your final shareable video.
What You’ll Learn in this AI Video Course?
How to access Kling and choose the right plan
How to generate images and animations using text prompts
How to animate still images
How to use lip sync, native speech and AI sound effects to enhance your videos
How to restyle, inpaint, and upscale your content
How to add, swap, and delete video elements in multi-layered scenes
How to build talking avatars with voice, movement, and expression
How to create complete multishot videos from a single text prompt
Who This Kling AI Video Course is For?
This Kling AI course is perfect for:
Creators and marketers who want to produce stunning videos without editing software
Content teams and solopreneurs seeking faster video creation workflows
Educators, storytellers, and filmmakers exploring the next wave of AI video tools
Anyone curious about text-to-video and the future of content creation
Requirements
No prior video editing or design experience needed.
All you need is:
A computer with internet access
A Kling AI account (we’ll guide you through this setup)
A curious, creative mindset ready to explore what’s possible with AI
Who teaches this AI Video course?
This course is led by Philip Ebiner, a top-rated creative instructor with over a million students worldwide, and Dan, an AI video expert and founder of AI Video School.
Together, we’ve designed this course to get you up and running fast—with no fluff and all the practical steps you need to create your first AI video with Kling.
Let’s Start Creating
AI video is changing how we tell stories, build brands, and express ideas. Kling is one of the most exciting tools available today— and this course gives you the roadmap to use it effectively.
Watch the free preview and enroll now to get lifetime access, future updates, and all the tools you need to start creating incredible AI videos with Kling.