How to Start Making AI Videos In 2026 (Beginner to Advanced)
At a glance
- Length
- 13 min
- Channel
- Youri van Hofwegen
- Video from
- Aug 2026
- Rating
- ⭐⭐ Great video · 2/2
- Best for
- Content creators wanting efficient, scalable AI video workflows
What this video answers
- What is the first step if I've never used Higgsfield?
- Why does planning matter more than the prompt itself?
- How do I keep characters looking the same across multiple videos?
- Can I create custom locations instead of using defaults?
- What does "fully directed scenes" mean in the video?
Overview of Cinematic AI Video Creation in Higgsfield
This tutorial walks through the complete process of making AI-generated videos using Higgsfield, a platform designed for creators who want to produce cinematic-quality content without traditional filming. The video covers the entire spectrum from absolute beginners through to advanced techniques, demonstrating how to transform a simple text prompt into a fully directed scene with consistent characters, custom environments, and professional-looking results.
The core message is that thoughtful planning and structured workflows produce better videos while consuming fewer generation credits—meaning cost efficiency alongside quality. Rather than treating AI video creation as a one-prompt-and-done process, the tutorial emphasizes building reusable assets and maintaining visual consistency across multiple scenes.
Key Strengths and Workflow Stages
- Progressive complexity: starts with basic prompts and scales toward storyboards, character design, and location creation
- Credit efficiency focus: teaches how to achieve higher quality with fewer resources, a practical concern for anyone paying per generation
- Reusable asset strategy: shows how to build and maintain consistent characters and custom locations across multiple videos
- Practical comparison approach: stages are shown side-by-side so viewers understand the quality difference each step brings
- Planning over prompt-writing: emphasizes that structure and preparation matter more than crafting the perfect single sentence
- Scene direction control: demonstrates how to guide AI generation for consistent, directed output rather than random results

Who Should Learn This AI Video Method
This tutorial suits creators, marketers, and content producers who want to generate video content quickly without owning expensive camera gear or hiring production crews. It's especially valuable for people already paying attention to their resource usage—whether that's time, money, or generation credits—since the efficiency angle runs throughout.
Beginners will find an accessible entry point with simple prompts, while experienced creators will benefit from the advanced techniques around character consistency and scene control. If you're exploring AI video tools to scale your content pipeline or test concepts before investing in live-action production, this workflow gives you a structured path. The verdict: solid practical guidance regardless of your current skill level.
Frequently Asked Questions About AI Video Generation
What is the first step if I've never used Higgsfield?
The video starts with a simple prompt—no special preparation needed. You write a description of what you want to see, and Higgsfield generates a video from it. This is the foundation before moving to more complex techniques.
Why does planning matter more than the prompt itself?
Planning means organizing your shots, characters, and locations before generation. The video demonstrates that a thoughtful structure produces more consistent and higher-quality results than trying to pack everything into a single detailed prompt.
How do I keep characters looking the same across multiple videos?
The tutorial shows how to create reusable character profiles within Higgsfield, then apply them consistently across scenes. This removes the randomness of regenerating characters and ensures visual continuity in your project.
Can I create custom locations instead of using defaults?
Yes. The video walks through building custom locations that you can save and reuse across multiple scenes, similar to the character approach. This gives you control over the visual environment and maintains consistency.
What does "fully directed scenes" mean in the video?
Rather than letting the AI generate whatever it wants, fully directed scenes mean you specify camera angles, character positions, timing, and other details to guide the output. The video shows this produces predictable, professional-looking results.

Key Terms
- Storyboards
- A planned sequence of scenes showing how your video will flow, shot by shot.
- Consistent characters
- AI-generated people or figures that look the same across multiple video scenes.
- Custom locations
- Unique, user-defined environments that can be saved and reused in different videos.
- Generation credits
- The resource units consumed each time you ask Higgsfield to render a video.
- Scene direction
- Specifying camera angles, character placement, and other technical details to control AI output.
Sources: Storyboards · Consistent characters · Custom locations · Generation credits · Scene direction — definitions cross-referenced with Wikipedia
Video by Youri van Hofwegen on YouTube. If you enjoyed it, please subscribe to their channel and show your support for the great video.
Description
Create the BEST AI Videos using Higgsfield 👉 https://youricreates.com/Higgsfield
In this video, I show you the complete workflow for creating cinematic AI videos in Higgsfield, starting with a simple prompt and building up to storyboards, reusable characters, custom locations, and fully directed scenes with consistent results. Along the way I compare each stage, explain why planning matters more than prompt writing alone, and show how to create higher-quality videos while using fewer generation credits.
✅ Generate top tier AI video prompts for free: https://www.videoprompt.studio/
for inquiries: yvh (at) youripartnerships.com
Video transcript Accessibility
A full written transcript of this video, provided for accessibility. Select any timestamp to jump the video to that moment.
I've been making AI videos for a while now, and the biggest thing I've learned is that to make a cinematic output, most people focus on the wrong thing. They think it's about writing the perfect prompt or finding the best model, but that's barely where the difference comes from. So, in this video, I'll walk you
through the complete [music] process for making a simple video all the way to a cinematic shot, so you can finally master AI video making. We're starting at the entry level, which is what most [music] people do, to show you exactly how much control you hand over to the AI without even realizing it. For this
whole video, I'll be working inside Higgsfield, which is an all-in-one platform that keeps every tool we need in one place. So, if you want to follow along, I've left a link for it in the description [music] below. Once you log in, you'll notice everything's organized at the top navigation bar. And to make a
video, you just click on video, which takes you straight into the video generation workspace. From the model selector, I'll pick Seedance 2.0, because right now, it's one of the best video models out there when it comes to cinematic outputs. So, the difference in quality will come down to how I control
the tools and not to the model. The whole idea at this stage is that you don't do anything technical, you just type out what you want in plain words. So, I'll set the duration to 10 seconds, the resolution to 1080p, and the aspect ratio to 16x9, then write one simple prompt for a motorcycle rider racing
across a desert canyon road. That's the idea I'll use for all four stages, so we can really test the difference in outputs based on the workflow we'll use. The settings, the model, and the core idea will all stay the same. And for how simple this was, it looks good, and technically, it's doing
exactly what I asked for. The bike is racing across the canyon with the dust kicking up exactly [music] like I described, and it even added its own engine sound and some camera movement. If you want to experiment with an idea or need a result fast, this will get your job done. But every choice in that
video, including the camera, the sound, [music] and how it looks in general, all came down to the model. None of it could capture my exact idea because I didn't give Seederns enough context to cover everything. At this stage, that's completely fine. You can run a couple of these simple prompts, pick the strongest
one, you've got a real video in almost no time, which is exactly what you want when you're just testing an idea. But the moment you want something specific instead of whatever the model gives you, writing a simple prompt isn't enough. You need something much stronger. But trying to write a detailed prompt
yourself can be very difficult, and it's easy to leave out important details. That's why you can just get an AI chatbot to do it for you in just a few minutes. So, I'll open up Claude and get it to write the prompt for me. I'll ask for a more detailed, polished video prompt for that same motorcycle rider
racing across the desert canyon, this time on a timed rally stage. And what it hands back is on a whole different level from my one line. It's a way more AI-friendly prompt with a full three-shot sequence. Seederns 2 is built for multi-shot generation from a single prompt, so this works great. It includes
a drone shot high above the canyon, the camera following the rider from the side, and even a bit of dialogue. It works so much better because a prompt like this gives Seederns actual directions to follow and real camera movements, so there's way less for it to guess at on its own. Then, I'll head
back to the video generation workspace with the same settings as before and paste in the prompt from Claude. >> [music] >> Right away, it's a completely different video. The shots cut the way Claude laid them out, and the camera moves actually feel deliberate. So, the whole thing
already looks a lot closer to a real scene from a film. And that jump has nothing to do with the model. It's the exact same Seederns 2.0 on the exact same settings. The only thing I changed was giving it a stronger prompt to work from. but it's still not perfect. The dialogue line I asked for never got
generated, but it's far closer to what I had in mind compared to the first video. So, from now on, a detailed prompt is a necessary step. You can do it using any chatbot you prefer on the free plan, and in just a few minutes, you will get a far better result than you would by writing it on your own. But, a prompt
can only ever describe what you want. It cannot actually show the model any visuals. So, without any references, this is as far as you can get. But, there's actually a way to plan out your entire video before animating a single frame. To do this, we need to make a storyboard, which is basically a single
image that maps out every shot of the video in one go. So, the model can see the whole thing laid out instead of relying on just a prompt. To make it, click on image and pick GPT image 2 as the model. For a storyboard, we basically want multiple images in one. While they all maintain maximum realism,
and GPT image 2 is currently the best model for holding all of that together. So, a prompt for a three-panel storyboard entire scene: a drone shot high above the canyon, a close-up of the rider leaning over the bars, and a rear shot with the dust kicking up behind him. And the panels come back looking
great. Consistency across all three panels holds up, and the lighting is realistic. But, there is one thing I don't like. Currently, the road came out as smooth paved asphalt with lane markings, which doesn't really match the whole scenery. So, instead of regenerating the whole storyboard
[music] and leaving it up to chance again, I'll just edit that specific thing. So, I'll switch over to Cinema Studio from the top navigation bar. Once you're in, make sure you're on image mode, drop the storyboard in as a reference, and tell it to fix only panels one and three, turning the road
into an unpaved dirt one, and keeping everything else exactly as it is. This time, the dirt road matches the scene while maintaining everything else exactly the same. This is the storyboard [music] we're going to use for our video. Now, to animate it, I'll ask Claude for a video prompt, but this
time, I'll also upload the storyboard, so it has the exact context of how the scene is supposed to look like. I'll also point out the duration of the video I want to make so we can plan everything accordingly. And again, it comes back with a detailed prompt that goes through three different shots that match our
storyboard. Then, still inside Cinema Studio, I'll switch over to video mode. And the first time you use an image in here, it makes you run a quick eligibility check. So, go to image generation and check the storyboard, which is just HeyGen Field confirming the image is allowed before you animate
it. From there, I'll set the storyboard as my reference, the genre to action, the duration to 15 seconds, the resolution to 1080p, and the aspect ratio to 16 by 9, and turn the audio on. As you can see, Cinema Studio gives us a lot more to work with compared to plain video generation, and that control is
what takes the final result from a cool video to something that feels directed. So, I'll paste in the prompt from Claude and generate. The video follows the storyboard almost exactly. Every shot in the order I
planned it, with all the elements staying consistent all the whole way through. And the reason this matters so much is to avoid burning through your credits on bad results. Editing an existing image is a lot cheaper than redoing an entire video after you've animated it. So, not only do you ensure
that your output will be exactly like you imagined, but you end up saving money and time. So, by planning everything ahead, the video comes out right on the very first generation. Everything we've done up till now are the stages that take you from beginner to pro. But to actually master AI video
making, you need to learn to control every single element in your video, just like a real director. So, this last stage is where you stop letting the model pick anything on its own and take control of everything. The plan here is to make a full 30-second cinematic split into two 15-second shots. And to keep it
all consistent, I'll build my own reusable character and location. I'll start back in Cinema Studio 3.5 inside image mode, set the model to auto, and this time, instead of a random rider, I'll make a character sheet of myself using a collage of my own pictures. A collage like this works way better than
a single image because the model will be able to understand what I look like from every single angle and keep my facial identity intact. So, I'll prompt for a character sheet of myself with a full shot of me next to the bike and a close-up of my face, all on a neutral background, so there's nothing to
confuse the model later. We get back this image, and my face looks exactly the same. It changed my outfit to match the scene we're making, and with this step, I even get to control how the bike looks like. If you want to make any changes, you can edit it the same way we fixed the storyboard, but I'm happy with
this one, and from now on, I can drop that exact version of me into any shot and stay completely consistent. Then, I'll switch the model over to cinematic locations, which is Hicksfield's model built specifically for realistic environments, and prompt a desert canyon with a dirt road. This is our
environment, and it looks pretty good. The previous stages generated the canyon from scratch, but this time, we know exactly how it's going to look. And though it might seem like a small step, when you combine it with everything else, this process starts to feel a lot more like directing a video than just
generating one. Now, I've got both a character and the setting, so I'll head back to the image generation workspace with GPT Image 2 and build a six-panel storyboard this time because I'm planning a full 30 seconds as two separate shots with three panels each. And because both shots pull from that
same saved character, location, and storyboard, they'll stay consistent with each other without me having to chain one off the other. This is our storyboard. All our elements have remained consistent. So, with the whole thing mapped out, I'll animate it one shot at a time. For the first shot, I'll
ask Claude for a video prompt based on the top three panels, and it gives me the tense opening, the drone shot, the swerve around the loose rocks, and the jump over the gap. Then, in Cinema Studio, I'll switch to video mode and run the same eligibility check for my assets. And instead of just pasting in
the prompt, I'll use the settings Cinema Studio provides. There's no other tool right now that lets you control a video this much, and every single setting affects your final output. I'll keep the genre on action, 15 seconds, 1080p, 16 by 9 with the audio on. This time, I'll also click on the smile mark and set the
emotion to vigilance, so my face matches the tension of the scene. Then, I'll paste in the cloud prompt and generate. This looks exactly like the storyboard, [music] with my face holding the whole way through, and that tense, vigilant
emotion has carried through to the final video. For the second shot, I'll go back to cloud and ask for a prompt for the bottom row. And then, back in Cinema Studio, I'll keep most of the settings the same, but I'll switch the emotion to serenity, since the scene is coming to an end, and generate.
>> Copy. Stage clear. >> And again, each scene comes out exactly as planned. Every single element, the bike, the location, and even myself from behind, look consistent, and the dialogue came through this time. Then, I'll take both finished shots into a video editor. I'm using CapCut for this,
but you can use any editor you prefer. I'll just arrange my videos in order, and they should feel as one continuous shot, because they've been precisely built with the same elements. Let's look at our final result.
>> Copy, stage clear. >> Compare this to the first video we made and the difference should be clear. And the whole reason for all this control isn't just quality, it's efficiency. Because saving your character and location and planning every shot is what lets you get exactly what you pictured
in your own style without burning through credits to get there. It's also worth mentioning that Higgsfield is about to release SeeDrones 2.5, which is set to be the best model on the platform yet. So, everything you just saw will only get better from here. So, the same idea that started as one plain sentence
is now a full cinematic scene with the same character in every shot. And that's the whole path from your first AI video to actually directing them like you're on a real production set. Once you master each one of these stages, you'll be left with the most efficient workflow that produces high-quality videos
without you needing to burn through all your credits. So, if you want to get started and make your own cinematic AI videos today, use the link in the description to sign up to Higgsfield. Thanks for watching and I'll see you in the next one.
How videos are chosen here
Every video on Helicopterstour.com is hand-picked and reviewed by Justin — nothing is added automatically. Each one gets an original written guide and an honest rating: ⭐ 1 out of 2 means a good video worth your time, and ⭐⭐ 2 out of 2 means a great one we would recommend to anyone. The videos belong to their creators — every page links back to the original channel so you can subscribe and support them.
