Heidelberg AICurriculum
Track 5 · Intermediate
5.3

AI video generation

Turn a sentence — or a photo — into a video clip

7 lessons 2026-08-06 AI-generated

1Overview

A fast-moving split between proprietary frontier models you prompt and pay per second (Sora, Veo, Runway, Kling) and open-source models you self-host on a rented GPU for free (Wan, HunyuanVideo, LTX-Video). Clip length, audio, and resolution all vary sharply — no vendor has solved long-form video yet, so the goal is to understand the trade-offs rather than memorize a winner.

Text-to-video and image-to-video have split into two camps: proprietary frontier models (Sora, Veo, Runway, Kling) that you pay for by the second or by subscription and just prompt, and open-source models (Wan, HunyuanVideo, LTX-Video) that are free to use but need you to rent or own a serious GPU. Clip length, audio support, and cost per second all vary sharply between them — no proprietary model reliably produces long clips yet, and audio is inconsistent across the board. This space changes monthly, so the goal is to learn the category and the trade-offs, not to memorize one vendor as "the" answer. → Quick pick: cinematic quality → Sora; fast with native audio → Veo; pro VFX control → Runway; cheap with good physics → Kling; free on your own GPU → Wan, HunyuanVideo or LTX-Video.

1.1After this chapter you can
Generate a short video clip from a text or image prompt using a proprietary frontier model
Know which proprietary model to reach for by need: cinematic (Sora), fast/cheap (Veo), pro control (Runway), physics (Kling)
Understand the open-source alternative — what Wan/HunyuanVideo/LTX-Video need in GPU hardware, and why that is worth it at volume
Recognize the current limits: short clips, inconsistent audio, and a market that reshuffles every few months
1.2Which model gives cinematic quality?

Sora is the proprietary option that consistently delivers cinematic‑grade visuals, though you pay per second for its output.

1.3Can I generate video for free?

Open‑source options like Wan, HunyuanVideo, and LTX‑Video run on your own rented GPU at no per‑second fee, but require you to manage the hardware yourself.

1.4Do any models support audio?

Veo offers native audio alongside video generation, while other services may provide limited or inconsistent sound support.

2Matrix 6 rows · 7 tools

sora
veo
runway
kling
wan
hunyuanvideo
ltx-video
Open source / self-host
no
no
no
no
yes
yes
yes
Runs locally
no
no
no
no
partial
partial
partial
Max clip length
60s
6-8s
~12s
10s
~5s
~5s
~24s
Native audio
no
yes
no
no
no
no
yes
Resolution
720-1080p
~1080p
~1080p
1080p
720-1080p
1280x720
768-1080p
Cost
$0.10-0.50/sec
~$0.05/clip
$12/mo+
~$0.84/10s
$0 · GPU
$0 · GPU
$0 · GPU

3Lessons 7

3.1 Generate a short cinematic video with Sora 2

Sora 2 is an AI model that turns text or images into high‑resolution videos with native audio.

Create and download a 6‑second 1080p video from a simple prompt using the Sora 2 web interface.

  1. Open the Sora 2 generator page (sora-2.ai).
  2. Enter a short descriptive prompt such as “a surfer rides a wave at sunset”.
  3. Select 1080p resolution, 16:9 aspect ratio, and set Duration to 6 seconds.
  4. Click the Generate button and wait for the preview to appear.
  5. When generation finishes, click Download to save the MP4 file.
  • You'll see A 6‑second 1080p video matching the prompt, with synchronized sound effects and dialogue if described.
  • Takeaway Understanding how Sora 2’s basic UI controls (prompt, resolution, duration) map directly to the characteristics of the output video.

3.2 Create a cinematic prompt for Sora 2

A concise cinematic prompt that splits into subject, camera movement, lighting and mood layers for Sora 2.

Create a sub‑50‑word prompt that clearly defines those four layers

  1. Open the Sora 2 interface and locate the prompt input field
  2. Enter the subject description followed by a comma
  3. Add the camera motion, then the lighting style, and finish with the mood separated by commas
  4. Count the words to stay under 50 and click the Submit button
  • You'll see A short video that matches the described scene, moves as instructed, and reflects the chosen visual mood
  • Takeaway Layering subject, camera, lighting and mood in a tight prompt gives the model clear direction
  • Check Which four layers should you include in a cinematic AI video prompt and how does word count affect the result?

3.3 Create a multi‑scene storyboard video with Sora 2

Sora 2 Storyboard is a feature that lets you define several timed scenes in one prompt.

Produce a 25‑second video composed of three distinct scenes using the storyboard timing option.

  1. Return to the Sora 2 generator page and click the “Storyboard” tab.
  2. Write a combined prompt with scene separators, e.g., "Scene 1: mountain hikers reach a summit – 8s; Scene 2: they set up a campfire – 9s; Scene 3: sunrise over the peaks – 8s".
  3. Set the total duration to 25 seconds and keep resolution at 1080p.
  4. Press Generate and monitor the real‑time preview as each scene renders.
  5. After completion, download the single 25‑second video file.
  • You'll see A seamless 25‑second clip that transitions through the three described scenes with consistent characters and physics.
  • Takeaway Storyboard mode lets you orchestrate longer narratives by chaining multiple prompts into a single coherent video.

3.4 Generate a video with multiple cuts using Sora 2

A single prompt that tells Sora 2 where to place “cut to” commands for automatic scene transitions.

Generate a clip containing at least two automatic cuts between shots

  1. List the required shots, e.g., a character walking, a close‑up of an object, and a wide room view
  2. Compose a prompt using the pattern: subject + action + "cut to" + next subject + action + environment details
  3. Paste the completed prompt into the Prompt field in Sora 2
  4. Click Generate to start video creation
  5. Play back the result and confirm that cuts occur at the intended points
  • You'll see A video where the camera jumps from one described shot to the next with smooth transitions
  • Takeaway Embedding “cut to” lets the model handle sequencing, removing a separate editing step
  • Check What two words in the prompt signal Sora 2 to create automatic cuts instead of requiring manual editing?

3.5 Refine a Sora 2 video by adjusting one element at a time

A workflow that uses Sora 2’s draft editor to tweak one element—lighting, camera angle or pacing—without restarting the whole generation.

Iteratively improve a clip until its mood and framing match your vision

  1. Open the first video from Drafts
  2. Click Edit, replace the targeted phrase in the prompt (e.g., change “harsh daylight” to “soft golden‑hour light”)
  3. Press Regenerate to create a new version with the updated element
  • You'll see A series of draft versions ending with a clip whose lighting, camera angle and pacing align with your original vision
  • Takeaway Targeted prompt tweaks converge faster than rewriting the whole description, conserving generation credits
  • Check Why does changing just one phrase in the prompt and regenerating produce better results than rewriting the entire description?

3.6 Add native audio to a Sora 2 video using sound effect mode

Sora 2 can generate synchronized dialogue and sound effects when the prompt describes audio cues.

Enhance an existing short video with realistic background sounds by specifying them in the prompt.

  1. Open a previously generated 6‑second video file on the Sora 2 page (use the “Create Video” section).
  2. In the prompt box, add audio descriptors such as “with roaring ocean waves and distant gulls”.
  3. Leave resolution and duration unchanged, but set Sound Effect to “On”.
  4. Click Generate to re‑render the clip with the added audio layer.
  5. Download the updated video and play it back to confirm the sound matches the description.
  • You'll see The same visual footage now includes synchronized ocean wave sounds and gull calls matching the prompt.
  • Takeaway Explicit audio cues in the prompt activate Sora 2’s native sound generation, allowing you to control both visuals and soundtrack together.

3.7 Export a high‑quality 4K video from Sora 2 for commercial use

Sora 2 offers HD/4K export options that produce watermark‑free, commercially usable files.

Generate a 10‑second 4K video and verify it is ready for commercial distribution.

  1. Navigate to the Sora 2 generator and choose the “Create Video” tab.
  2. Enter a detailed prompt such as “a futuristic city skyline at night, neon lights flickering, with flying cars”.
  3. Select Resolution 4K (if available), Aspect Ratio 16:9, and set Duration to 10 seconds.
  4. Ensure the option for no watermark is active (the platform states videos are delivered without watermarks).
  5. Click Generate, wait for completion, then click Download and confirm the file’s resolution in your media player.
  • You'll see A 10‑second 4K video with cinematic lighting and motion that plays back without any watermark.
  • Takeaway Choosing higher resolution and confirming watermark settings ensures the output meets professional standards for commercial projects.

4You’ll know it worked 23 checkable outcomes in this chapter

  • The generated prompt includes explicit timecodes, camera directions, and scene breakdowns that match the documentation’s structure.
  • The generated video maintains the exact character design, props, and environmental details from the uploaded references.
  • The video displays consistent shadows, highlights, and color temperature matching the described light source
  • The video plays continuously without visible seams, jumps, or directional logic breaks
  • The output video features the exact product model, color, or architectural details from the uploaded reference image
  • The revised generation matches your exact creative vision after 2 to 3 focused adjustments
  • The output video plays with matching visuals, character actions, and background music as described.
  • The cameo icon appears in the prompt toolbar and populates your video with your likeness

23 outcomes in all — one per recipe below.

5FAQ, Tips & How-to 27

one problem, one solution, one action
FAQ Everyone

How do I add my own likeness as a cameo in Sora videos?

Tap the plus icon, choose Add Yours, and upload a clear photo of yourself. Set who can see the cameo (just you, friends, or everyone) and then reference your name in the text prompt to place the avatar in any scene.

AI-generated
FAQ Everyone

What is the easiest way to get access to Sora 2?

You need a ChatGPT Plus or Pro subscription and an invite code, which can be found on the official OpenAI Discord, X, or shared by the community. Install the iOS app or visit sora.com, link your subscription, and enter the code.

AI-generated
FAQ Everyone

How can I improve a video that didn’t turn out right the first time?

Open the draft in the drafts folder, tap Edit to reopen the prompt, and change one element such as lighting, camera angle, or soundtrack. Regenerate the clip; this iterative tweak saves generation credits and aligns the output with your vision.

AI-generated
FAQ Everyone

What should I include in my prompt to get a professional‑looking video?

Start with the video's purpose (e.g., "TV ad for a jewelry brand"), then add explicit camera directions (like "cinematic close-up"), lighting details (such as "soft natural daylight"), and any material or texture cues. Keeping the prompt under 50 words helps Sora focus on the key visual cues.

AI-generated
How-to ChatGPT Everyone

Only have a short idea and images

Uploading OpenAI’s official Sora 2 documentation into ChatGPT trains the LLM on the exact terminology, formatting, and structural expectations of the video model. ChatGPT then acts as a prompt architect, converting brief concepts and reference images into highly detailed, timecoded scripts that specify camera movement, lighting, pacing, and mood. This bridges the gap between vague text inputs and the model’s actual rendering logic.

Artlist ↗ Lesson → AI-generated
How-to Everyone

Awkward dismounts or static cameras in AI video

AI video models rarely produce perfect results on the first generation due to complex physics and motion synthesis. By treating ChatGPT as a cinematographer that expands your directorial notes, you can isolate specific flaws like awkward dismounts or static cameras and request precise shot-level fixes. Regenerating after each refinement cycle or simply re-running the prompt often yields the final polished clip.

Artlist ↗ Lesson → AI-generated
How-to Everyone

Text prompts keep changing characters and settings

Text-only prompts often cause AI video models to hallucinate inconsistent characters, props, or environments. Uploading reference images alongside your prompt gives the model concrete visual anchors, allowing it to extract accurate details for lighting, texture, and composition. This dramatically reduces random variations and keeps the generated video aligned with your original vision.

Artlist ↗ Lesson → AI-generated
How-to Everyone

Not sure what style my video should have

Telling Sora the intended use (e.g., TV ad, YouTube explainer) gives the model a clear creative boundary, focusing its attention on relevant visual and tonal cues. This prevents vague outputs and aligns the generation with professional standards. The model uses the stated purpose to infer pacing, polish level, and compositional priorities.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Unsure what view the AI will use

Sora relies heavily on explicit camera direction to determine composition. Describing angles (close-up, wide shot, tracking shot) directly dictates the viewer's visual experience and emotional engagement, replacing guesswork. The model maps these directional cues to specific framing and depth-of-field behaviors.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Lighting feels flat in AI video prompts

Lighting is the primary driver of perceived realism and cinematic quality in AI video. Explicitly stating light sources, times of day, or quality (soft, harsh, directional) gives Sora a concrete reference for shading, reflections, and atmosphere. This reduces the model's reliance on default, often flat, lighting setups.

AI Master ↗ Lesson → AI-generated
How-to Everyone

5‑second clip limit stops long scenes

Sora's 5-second limit and predictive nature struggle with complex, multi-step actions in a single prompt. The Storyboard feature chains sequential prompts, maintaining character and environment consistency while maximizing action within each clip. This turns isolated generations into a continuous narrative flow.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Need a custom video transition

The Blend feature merges two separate videos using adjustable timing curves. The Transition mode acts like a manual keyframe system, letting users control exactly when and how one clip morphs into another. This bypasses rigid preset transitions and creates unique, user-directed morphs.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Need a looping background video

Sora's Loop feature extends a clip by blending its start and end frames. It works best on repetitive, non-directional motion because forward-moving objects create jarring discontinuities when merged. This makes it ideal for generating ambient backgrounds or motion graphics.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Videos look flat and unrealistic

Sora struggles with complex physics but excels at rendering surface properties. Explicitly naming materials grounds the scene and tricks the model into generating convincing light interactions and weight. This anchors the video in physical reality, making abstract or simple motions feel tangible.

AI Master ↗ Lesson → AI-generated
How-to Everyone

My AI videos turn out random and low quality

Sora 2 responds best to structured, cinematic prompts rather than vague descriptions. Using templates like scene plus camera movement plus style gives the model clear directional cues for lighting, framing, and pacing. This structure mimics film directing, resulting in predictable, high-quality outputs.

AI Master ↗ Lesson → AI-generated
How-to Everyone

AI video keeps showing generic versions of my product

AI video models often hallucinate generic versions of specific objects or places. Uploading a reference photo before generation anchors the model to the exact visual details of your subject. This dramatically improves accuracy for client work, product demos, and branded content.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Need a single video with built‑in scene changes

Instead of generating separate clips and editing them together, you can instruct the AI to produce a sequence with internal cuts. Using the phrase cut to between scene descriptions tells the model to transition between shots while maintaining continuity. This saves time and creates dynamic storytelling clips.

AI Master ↗ Lesson → AI-generated
How-to Everyone

When AI video glitches on splashing water or flowing fabric

AI video models still struggle with complex physics like splashing liquids, flowing fabric, or gravity-defying actions. Prompting for simple, static, or gently moving subjects prevents weird artifacts and unnatural movement. This strategy ensures reliable, watchable outputs without wasting generations.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Video is close but the mood’s off

First-generation AI videos are often 80 percent correct but miss the exact mood or framing. Instead of rewriting the entire prompt, tweak one variable at a time (lighting, camera angle, pacing) and regenerate. This iterative approach builds on successful outputs and rapidly converges on the desired result.

AI Master ↗ Lesson → AI-generated
How-to Everyone

Can't access the limited rollout app

Sora 2 is currently in a phased rollout that requires a ChatGPT Plus or Pro subscription plus an invite code. Users can obtain codes by monitoring the official OpenAI Discord, searching X, or checking community shares as OpenAI gradually expands access. This gatekeeping manages server load while gathering early creator feedback.

Kevin Stratvert ↗ Lesson → AI-generated
How-to Everyone

Want your own face and voice in AI videos

Sora 2 lets you train a temporary digital avatar of yourself using a few seconds of video and voice. You grant permission, set visibility (private, friends, or public), and then reference your name in prompts to insert yourself into any scene. The model maps your facial features and voice onto the generated character, enabling personalized video creation.

Kevin Stratvert ↗ Lesson → AI-generated
How-to Everyone

Need a video that follows my exact scene, actions and music

Sora 2 generates video and synchronized sound from a single text prompt. By explicitly describing the scene, action, camera movement, and background music in one prompt, the model aligns the visuals and audio track. This unified prompting approach reduces the need for separate editing tools and leverages the model’s physics and audio synthesis capabilities.

Kevin Stratvert ↗ Lesson → AI-generated
How-to Everyone

Want broadcast‑quality video and a bit more length

Pro subscribers unlock a higher-tier model that prioritizes visual quality and longer clip duration over speed. The interface exposes advanced settings for resolution (high or standard) and duration (up to 15 seconds), giving creators more control over the final output. This tier is designed for users who need broadcast-ready quality or longer narrative clips.

Kevin Stratvert ↗ Lesson → AI-generated
How-to Everyone

Need to adjust a generated video's music, angle or style

After generation, videos land in the drafts folder where you can reopen the prompt box to make targeted edits. Instead of regenerating from scratch, you can tweak specific elements like the soundtrack, camera angle, or visual style by modifying the prompt. This iterative workflow saves time and helps fine-tune the output to match your vision.

Kevin Stratvert ↗ Lesson → AI-generated
How-to Everyone

Want your own face in AI‑generated videos

Sora 2 lets you upload a photo to create a digital cameo that the model can animate in new scenes. The app processes the image into a recognizable avatar, which you can then insert directly into text prompts to control who appears in the generated video.

CNET ↗ Lesson → AI-generated
How-to Everyone

Instead of starting from scratch, Sora 2's main feed includes a Pick a Mood filter that lets you type specific vibes or genres to surface existing AI videos. Swiping left on any video reveals alternative prompts or variations used to recreate it, providing instant inspiration for your own generations.

CNET ↗ Lesson → AI-generated
How-to Everyone

My AI video isn’t perfect

Initial generations often need tweaking. Sora 2 saves outputs to a drafts folder where you can reopen the prompt, modify details, and regenerate without losing the original concept. This iterative loop improves alignment with your vision and saves generation credits.

CNET ↗ Lesson → AI-generated

The same set on /recipes, filtered by tool and role.

6Videos 4

7FAQ 34

How can I keep the appearance of a specific product or location accurate in my AI video?

Upload a clear, well‑lit reference photo of the exact product or place and attach it to the reference image field before generating. The model uses that image as an anchor, reducing generic hallucinations and matching the visual details you need.

What’s the best way to generate a video with several cuts without editing separate clips?

Use multi‑cut prompting by describing each shot and inserting “cut to” between them (e.g., "person walks forward, cut to close‑up of hand, cut to wide room view"). The model then creates internal transitions, giving you a single clip with built‑in scene changes.

Why do some AI videos show weird splashes or unrealistic motion, and how can I prevent that?

Complex physics like liquid splashing or heavy wind often cause artifacts. Follow the physics‑safe strategy: replace those actions with simple, static or gently moving subjects (e.g., a bottle on a table instead of pouring wine) and use gentle camera moves to keep the output clean.

If my first generated video is close but not perfect, how do I improve it without starting over?

Apply the remix‑and‑iterate workflow: identify one element that’s off (like lighting or camera angle), change only that part of the prompt, and regenerate. Repeat tweaking single variables until the video matches your desired mood and framing.

Can I add my own likeness to a generated video, and what steps are required?

Yes, Sora 2 lets you create a cameo avatar by uploading a photo (or short video) of yourself, setting visibility permissions, and then referencing your name in the text prompt. The model animates that digital version within any scene you describe.

How can I keep the product look accurate in an AI‑generated video?

Upload a clear, well‑lit reference photo of the exact product and attach it to the reference image field before you generate. The model uses that image as an anchor, so the resulting clip matches the real product’s details instead of hallucinating a generic version.

What kind of prompt structure gives the best results with Sora 2?

Use a cinematic formula that lists the scene, camera movement, and style in under 50 words. For example, start with the camera direction (like “wide tracking shot”), then describe the subject’s action, followed by the visual mood or lighting.

I’m getting weird liquid splashes and other artifacts—how can I avoid them?

Follow the physics‑safe strategy: choose simple, static or gently moving subjects instead of complex actions like pouring liquids or heavy wind. Specify gentle motions such as a slow rotation or soft drape to keep the output clean.

Can I change just the lighting or camera angle without redoing the whole video?

Yes, use the draft editor: open the generated clip in your drafts folder, tap Edit, modify only the part of the prompt that describes lighting or camera perspective, and regenerate. The rest of the video stays the same.

How do I create a multi‑shot sequence without stitching separate clips together?

Use multi‑cut prompting by writing a single prompt that includes “cut to” between scene descriptions. List each shot’s subject and action in order, and Sora 2 will generate internal transitions for you.

How can I keep the AI video looking exactly like my product or location?

Upload a clear, well‑lit reference photo and attach it in the reference image field before you generate. The model uses that image as an anchor, so the resulting clip matches the exact visual details of your subject.

What’s the best way to write prompts for high‑quality video clips?

Use structured cinematic formulas such as “scene + camera movement + style” and keep the prompt under 50 words. Fill in specific details like location, camera action, and mood; this gives the model clear directional cues for lighting, framing, and pacing.

I’m getting weird splashes or floating objects—how can I avoid those artifacts?

Follow the physics‑safe strategy: replace complex actions (like liquid splashing or heavy wind) with simple, static or gently moving subjects. Specify gentle camera moves such as slow rotation to keep motion natural and prevent unnatural artifacts.

Can I create a video that has multiple shots without editing separate clips?

Yes, use multi‑cut prompting by writing “subject + action + cut to + next subject + action…”. The phrase “cut to” tells the model to insert internal transitions, producing a single clip with built-in scene changes.

My first video is close but not perfect—how do I refine it without starting over?

Apply the remix and iterate workflow: identify one element that’s off (e.g., lighting or camera angle), change only that part of the prompt, and regenerate. Repeat tweaking single variables until the clip matches your desired mood and framing.

How can I add my own likeness to an AI‑generated video?

Upload a personal photo in Sora 2’s cameo feature; the app creates a digital avatar from it. Then reference your name or the cameo in the text prompt, set visibility permissions, and generate the video with you appearing in the scene.

What’s the best way to write a prompt for Sora 2 so the video looks professional?

Use a structured, cinematic formula that includes the scene, camera movement, and style, keeping it under 50 words. This gives clear direction for lighting, framing, and pacing, which helps the model produce high‑quality clips.

How can I make sure the AI uses my exact product image in the video?

Upload a clear, well‑lit reference photo in the reference image field before generating. The model anchors its visuals to that image, reducing hallucinations and keeping the product appearance consistent throughout the clip.

Can I create multiple shots without editing separate clips together?

Yes, use multi‑cut prompting by writing “cut to” between scene descriptions. This tells Sora 2 to insert internal transitions, producing a single video with built‑in cuts and continuous storytelling.

Why do some videos have weird motion like splashing liquids, and how can I avoid it?

Complex physics such as liquid splashes often cause artifacts. Prompt for simple, static or gently moving subjects—like a bottle on a table instead of pouring wine—to keep the motion natural and artifact‑free.

+ 14 more in the library.

8Glossary 78 terms

Show the 78 terms
Sora 2
Cameo
A digital avatar of your own likeness that you can insert into Sora videos by uploading a photo and granting permission.
Pick a Mood
A filter in the main feed where you type a vibe or genre to see AI videos matching that mood.
draft editor
The tool that lets you reopen a generated video’s prompt, make changes, and regenerate without losing the original draft.
Storyboard
A feature that chains multiple short prompts together to create a continuous multi‑scene video sequence.
Blend tool
A function that merges two videos using adjustable timing curves to control how one clip morphs into another.
Loop tool
A feature that extends a clip by blending its start and end frames to create a seamless repeating background video.
Multi‑Cut Prompting
A prompting style where you write ‘cut to’ between scene descriptions so Sora creates internal cuts within one video.
Physics‑Safe Prompting
A strategy of describing simple or static actions instead of complex physics to avoid visual artifacts in the output.
Cinematic Prompt Formulas
Structured templates that combine scene description, camera movement, and style to guide Sora toward professional‑looking clips.
Reference Image Integration
Uploading a photo alongside your prompt so Sora can use its visual details for more accurate video generation.
Google Veo 3
reference image field
The place in the video generation interface where you upload a photo to guide the AI’s visual output.
drafts folder
A storage area in Sora where generated videos are saved so you can reopen and edit them later.
Pick a Mood
A filter on Sora’s main feed that lets you type a vibe or genre to see example AI videos matching that mood.
cameo permissions
Settings that decide who can see the digital avatar of your likeness, such as only you, approved friends, or everyone.
invite code
A special alphanumeric key required to join Sora’s limited rollout when you have a ChatGPT Plus or Pro subscription.
ChatGPT Plus
A paid version of ChatGPT that provides higher usage limits and is needed to access Sora’s advanced features.
Multi-Cut Prompting
A way to tell the AI to create several shots with internal transitions in a single video by using the phrase “cut to” between scene descriptions.
Physics‑Safe Prompting
A strategy of avoiding complex actions like splashing liquids or heavy wind in prompts to prevent visual glitches in the generated video.
Blend feature
A tool in Sora that merges two videos into one by adjusting a timing curve for a smooth transition.
Loop tool
A function that extends a video by blending its start and end frames to create a seamless repeating clip, useful for background footage.
Runway Gen-4
reference image field
A place in the video generation interface where you upload a photo to guide the AI’s visual output.
drafts folder
A storage area that saves generated videos and lets you reopen and edit their prompts later.
Pick a Mood filter
A search option on Sora’s main feed that shows videos matching a typed vibe or genre.
Cameo permissions
Settings that decide who can see or use your uploaded likeness, such as only you, friends, or everyone.
invite code
A special alphanumeric key required to gain access to Sora during its limited rollout.
ChatGPT Plus or Pro subscription
A paid plan that unlocks higher‑tier features and is needed to use Sora 2 fully.
Blend feature
A tool that merges two videos into one by adjusting a timing curve for the transition.
Loop button
A control that creates a seamless repeat of a video by blending its start and end frames.
Storyboard button
An option that lets you chain several prompts together to build a longer narrative sequence.
cut to phrase
The words “cut to” used in a prompt to tell the AI to insert a scene transition between shots.
physics‑safe prompting strategy
A guideline to avoid describing complex motions like splashing liquids, which often cause visual errors.
Sora 2 app
The mobile or web application where you create, edit, and generate AI videos with Sora’s tools.
Kling
reference image field
The place in the Sora interface where you upload a photo to guide the AI’s visual output.
multi-cut prompting
A way of writing a prompt that tells the AI to create several shots with cuts inside one video generation.
physics-safe prompting
Choosing simple actions and motions in your prompt to avoid unrealistic effects like splashing liquids or impossible movements.
cameo avatar
A digital copy of your likeness that Sora can animate in generated videos after you upload a photo or short video of yourself.
Pick a Mood feed
A filter on the Sora main page where you type a vibe or genre to see example AI videos matching that mood.
draft editor
The tool inside Sora that lets you reopen a generated video’s prompt, edit details, and regenerate without losing the original version.
invite code
A special alphanumeric key required to create a Sora account during its limited rollout.
ChatGPT Plus
A paid subscription tier for ChatGPT that is needed to access Sora’s advanced video generation features.
Pro subscription
An upgraded plan (ChatGPT Pro) that unlocks higher‑quality, longer video generation and extra settings in Sora.
Blend feature
A Sora tool that merges two videos by adjusting a timing curve to create custom transitions.
Loop tool
A function in Sora that extends a clip by smoothly joining its start and end frames for seamless background loops.
Storyboard button
The interface control that lets you chain multiple short prompts together to build a longer, continuous video sequence.
Wan 2.2
reference image field
The place in the video generation interface where you upload a photo to guide the AI’s visual output.
cut to
A phrase used in prompts that tells the model to transition from one scene description to the next within the same video.
draft editor
The tool that lets you reopen and modify a previously generated video's prompt before regenerating it.
invite code
A short alphanumeric key required to gain access to Sora 2 during its limited rollout.
ChatGPT Plus
A paid subscription tier for ChatGPT that is needed to use Sora 2 and obtain an invite code.
cameo permissions
Settings that control who can see or use your uploaded personal avatar in generated videos (e.g., only you, friends, or everyone).
Pick a Mood filter
A search option on Sora’s main feed where you type a vibe or genre to see AI‑generated videos matching that mood.
Storyboard feature
A function that lets you chain several short prompts together to create a longer, continuous video sequence.
Blend feature
A tool for merging two separate video clips by adjusting timing curves to control the transition between them.
Loop tool
An option that extends a clip by smoothly joining its start and end frames, creating a seamless repeating video.
high resolution setting
A selection in Sora 2 Pro that sets the output video’s detail level to ‘High’ for better visual quality.
physics‑safe prompting
A strategy of avoiding complex actions like splashing liquids or heavy wind in prompts to prevent unnatural AI video artifacts.
HunyuanVideo
Cinematic Prompt Formulas
Pre‑made sentence structures that tell the AI what scene, camera movement and style to use, helping it create professional‑looking video clips.
Reference Image Integration
Uploading a photo alongside your text prompt so the AI can copy exact visual details from that image into the generated video.
Multi-Cut Prompting
Writing a single prompt that includes the phrase “cut to” to tell the AI to create several shots with transitions in one video.
Physics‑Safe Prompting Strategy
Choosing simple actions and gentle movements in your description to avoid AI glitches caused by complex physics like splashing water or flowing fabric.
Remix and Iterate Workflow
A step‑by‑step method of changing only one part of a prompt at a time (like lighting) and regenerating the video until it matches what you want.
Pick a Mood feed
A filter in the Sora app where you type a vibe or genre and see AI videos that match, letting you copy their prompts for inspiration.
draft editor
An interface inside Sora where saved video drafts can be opened, edited, and regenerated without losing the original prompt.
invite code
A short alphanumeric key you obtain from community channels that unlocks access to Sora during its limited rollout.
Cameo avatar
A digital version of your own face created from a photo or short video, which you can insert into AI‑generated scenes via text prompts.
Blend tool
A function that merges two separate videos by adjusting a timing curve so one clip smoothly morphs into the other.
Loop tool
A feature that extends a video by blending its start and end frames, creating a seamless loop ideal for background footage.
Storyboard feature
An option that lets you chain several short prompts together in sequence, producing a longer narrative made of multiple AI‑generated clips.
LTX-Video
reference image field
A place in the video generation interface where you upload a photo to guide the AI’s visual output.
drafts folder
A storage area that saves generated videos so you can reopen and edit their prompts later.
Pick a Mood filter
A tool on Sora’s main feed that lets you type a vibe or genre to see AI videos matching that mood.
Cameo
A digital avatar of yourself created from an uploaded photo, which the AI can insert into generated scenes.
invite code
A special alphanumeric key required, along with a ChatGPT Plus or Pro subscription, to access Sora 2 during its limited rollout.
Blend feature
A function that merges two videos by adjusting timing curves to create custom transitions between them.
Loop tool
A utility that extends a video by smoothly joining its start and end frames for seamless background loops.
Storyboard button
An interface element that lets you chain multiple prompts together to build longer, sequential video narratives.
cut to
A phrase used in a prompt to tell the AI to transition from one scene description to the next within the same clip.
Physics-Safe Prompting Strategy
A guideline advising you to avoid describing complex physical actions like splashing liquids, which the AI often renders poorly.

9See also

💬 Discuss this chapter

Ask, share, or report — over on the Heidelberg AI community forum.