Skill Deep-Dive 8 min read Last updated: May 2026

The AI Film Prompt That Gives Sora a Character, Not Just a Scene

Generic AI video prompts describe what happens. A cinematically calibrated prompt specifies what the camera sees, how the character holds their body, and what the light does to the scene — and those three things are the difference between footage and film.

SP
Founder, NovaKit
🎬
NovaKit Skill
Dialogue / Character Film Prompt — cinematically calibrated prompts for Sora, Runway, and Kling that produce footage with intentional visual grammar
Quick answer: Generic AI video prompts describe what happens. A cinematically calibrated prompt specifies what the camera sees, how the character holds their body, and what the light does to the scene — and those three things are the difference between footage and film.
In this guide

Dialogue / Character Film Prompt is a Claude AI skill — cinematically calibrated prompts for Sora, Runway, and Kling that produce footage with intentional visual grammar.

  1. Why Does AI Video Generation Produce Generic-Looking Footage?
  2. What Does Cinematic Calibration Change About the Prompt?
  3. What Does the Dialogue / Character Film Prompt Skill Produce?
  4. Narrative Prompt vs Cinematically Calibrated Prompt
  5. Is This Skill Right for Your Production?
  6. Frequently Asked Questions

You've written the prompt. A woman sits at a kitchen table in the early morning. She's thinking about something difficult. The light is coming through the window. The footage comes back and it is technically correct in every way — a woman, a table, morning light, a contemplative expression. It looks like the establishing shot in ten thousand other films. It looks like AI.

The prompt wasn't wrong. It just gave the model nothing to work with beyond the narrative situation. It didn't specify whether the camera is at table level looking up, or mounted high looking down — two framings that produce completely different power dynamics for the character. It didn't specify whether the morning light is hard and directional, casting sharp shadows across the table surface, or diffuse and overcast, flattening the scene into something quieter. It didn't say whether the character is absolutely still or whether her fingers move slightly against the table. These are not details — they're the decisions that distinguish a scene with a point of view from footage that happens to contain a person.

The Dialogue / Character Film Prompt skill is built around a single principle: AI video tools don't fail because the technology is weak. They fail because the prompt was written without scene grammar.

Why Does AI Video Generation Produce Generic-Looking Footage?

Generic AI footage is the direct output of generic prompts — descriptions of narrative situations rather than visual specifications. When a prompt tells an AI video tool what happens in a scene, the model fills every unspecified visual decision with its training distribution average: the most common camera height, the most common lighting setup, the most common expression for the emotional state described. That average is what generic footage looks like. It is the visual equivalent of a sentence that is correct but says nothing.

💡
The core problem

A film prompt is not a scene description. It is a camera specification, a lighting specification, and a character physical specification simultaneously. Every element the prompt leaves unspecified gets filled by the model's training average — which is the definition of generic. Specificity is not polish; it is the actual content of the prompt.

The problem compounds for dialogue and character scenes specifically. Action sequences can absorb some visual ambiguity — the camera follows the action and the action provides the structure. Dialogue scenes have no inherent action to follow. Without precise specification of which character holds frame, where the cut falls, how physically close the characters are to each other, and what the camera relationship between them expresses about their dynamic, the model produces footage of two people talking that could belong to any scene in any film. The dialogue is in the writing. The scene is in the visual grammar — and visual grammar requires visual specification.

That gap is exactly what the Dialogue Character Film Prompt skill for Claude was built to close.

What Does Cinematic Calibration Change About the Prompt?

Cinematic calibration replaces narrative description with visual specification across the five layers that determine what AI video footage actually looks like. Before writing a word of the prompt, the Dialogue / Character Film Prompt skill researches current cinematic references in the genre and visual register you're working in, then builds the prompt from those five layers outward.

Prompt layer Generic prompt (narrative) Calibrated prompt (visual specification)
Camera Not specified — model chooses "Handheld, medium close-up, slight push in, 50mm equivalent — camera stays with her face as she speaks"
Character body "She looks nervous" or "she seems upset" "Sits very still, weight shifted back, hands flat on the table — movement only in her eyes tracking left before she speaks"
Light "Morning light through the window" "Hard directional side light from frame left, warm 4500K, casting table shadow — her left side lit, right side in shadow"
Colour grade Not specified "Slightly desaturated, teal shadows, muted highlights — reference: A24 naturalistic drama palette"
Scene grammar "Two people having an argument" "Tight over-shoulder shot, camera axis 30 degrees off eyeline — the frame cuts off the other character at the shoulder, creating pressure without showing their reaction"

The calibrated prompt is not longer for the sake of detail — it's more specific about the things that actually determine the visual output. A model given precise camera, character, lighting, colour, and scene grammar specifications produces footage with intentional visual decisions. A model given a narrative situation produces footage where every visual decision was made by averaging its training data.

The prompt is the direction. What you don't specify gets decided for you — by the model's training average, which is the most generic version of whatever you asked for.

What Does the Dialogue / Character Film Prompt Skill Produce?

Three complete prompt variants per scene, each with a different camera approach, plus a master character reference document for visual consistency across multiple shots.

1
Cinematic reference research
Before writing, the skill identifies current cinematic references — directors, films, or visual movements — that match the scene's emotional grammar and visual register. These references anchor the prompt's lighting language, colour palette, camera proximity, and movement style. A scene calibrated to Kelly Reichardt's visual grammar produces different footage than the same scene calibrated to Wong Kar-wai's — and the difference is in the prompt's specificity, not in post-production.
2
Character visual language profile
From the character description and scene context, the skill derives the character's specific physical vocabulary: how they carry weight, the quality of their stillness or movement, where their eyes go, how they occupy space relative to another character. This profile is embedded in the prompt as physical specification rather than emotional instruction — "her right hand lifts slightly then stops" rather than "she's thinking about whether to reach out."
3
Three prompt variants — three camera approaches
The primary prompt uses the camera approach that best serves the scene's emotional grammar. The second variant offers an alternative approach — tighter or wider, static versus moving, different axis — for comparison or for the editor who wants options. The third variant specifies a different lighting setup for the same scene. All three are formatted for the target tool (Sora, Runway, Kling, or equivalent) with the technical parameters that tool responds to most reliably.
4
Master character reference
A consolidated character visual reference document — physical description, movement vocabulary, lighting relationship, colour response — formatted for reuse across multiple shots and scenes. Paste it into subsequent prompts to maintain visual consistency across a sequence. Without a character reference, each new prompt risks generating a slightly different version of the character, which breaks continuity across shots.
NovaKit Skill
Dialogue / Character Film Prompt — three variants, character reference, cinematically calibrated
Works inside Claude. Three prompt variants per scene plus a master character reference for shot continuity. Formatted for Sora, Runway, and Kling.
See the skill from $15 · instant download

Narrative Prompt vs Cinematically Calibrated Prompt

Same scene — a woman in her late thirties, alone in a kitchen, early morning. She's received news she hasn't processed yet. Two prompts; one written from narrative, one from visual specification.

Generic narrative prompt
A woman in her late thirties sits alone at a kitchen table early in the morning. She has just received some difficult news and is sitting quietly, looking thoughtful and a little sad. The morning light comes through the window. She seems lost in thought. Cinematic, emotional, realistic style.
✓ Cinematically calibrated prompt
35mm film grain. Handheld medium close-up, camera at table height, slight drift — no push. 50mm equivalent. Woman, late 30s, dark hair, sits completely still — hands flat on table surface, fingers spread. Eyes fixed on a point off-frame right. Hard directional morning light from a high window, frame left — warm 4800K, sharp shadow across the right half of her face and down her neck. Teal shadow base, slightly desaturated highlights. No music. Room sound only. She breathes in. Does not move. 8 seconds. Reference: Chloé Zhao, naturalistic drama, extended stillness.

The narrative prompt gives the model a situation. The calibrated prompt gives it a decision for every visual element that determines what the footage looks like: camera height (table level, looking slightly up), camera movement (handheld drift — neither static nor active), character physiology (completely still, fingers spread — specific physical choices), lighting (hard, directional, exact colour temperature, which side of the face), colour grade (teal shadows, desaturated highlights), duration (8 seconds), and a cinematic reference that anchors the whole frame in a specific visual tradition. The model still makes the final image. But its decisions are now constrained by intention rather than left to the training distribution average.


Is This Skill Right for Your Production?

The skill delivers the most value for character-driven scenes where the emotional content lives in physical specificity — the stillness, the eye movement, the weight — rather than in action. Dialogue scenes, contemplative moments, and character-establishing shots are where generic prompts fail most visibly and where calibrated prompts produce the largest improvement in output quality.

The next piece most people tackle from here is a podcast episode script with natural conversation flow. If you're working across the full Video & Pod workflow, the Video & Pod bundle covers everything in one place.

✓ This is for you if
You're generating footage for narrative film, short film, or character-driven content
Your AI video output consistently looks generic despite technically correct prompts
You need visual consistency across multiple shots of the same character
You're a screenwriter or director using AI tools in pre-visualisation or production
You're producing content for Sora, Runway, Kling, or equivalent generation tools
✗ This isn't for you if
You need abstract or non-representational AI video — this skill is character and scene specific
Your primary use case is product demonstration or explainer video without characters
You're generating social media loops or background footage without narrative content

Frequently Asked Questions

How do you write a good AI film prompt for Sora or Runway?
Effective AI film prompts specify what the camera sees rather than what happens in the scene. Include lens choice and approximate focal length, character physical specifics — posture, movement quality, eye direction — lighting direction and colour temperature, and a cinematic reference that anchors the visual language. Generic action descriptions give the model nothing to work with beyond narrative situation, producing footage where every visual decision defaults to the training distribution average.
Why does AI video generation produce generic-looking footage?
Because most prompts describe narrative action rather than visual specification. Telling a video generation tool "a woman walks into a room and looks around nervously" gives it no information about lens choice, character physicality, lighting grammar, or visual tone. The model fills every unspecified gap with its training distribution average — the most common choice for each unspecified element — which is what generic footage looks like.
What is character visual language in AI film prompts?
Character visual language is the specific physical vocabulary that distinguishes a character's presence on screen: how they carry weight, the quality of their movement, where their eyes track, how they occupy space relative to other elements in the frame. "She looks nervous" and "she stands very still, weight shifted to one foot, eyes tracking the exit door, jaw set" produce completely different footage from the same AI tool — because the second gives the model actual physical decisions to render.
What should a Sora or Runway film prompt include?
At minimum: camera position and movement (static, push, handheld drift), lens type and approximate focal length, character physical description and specific movement quality, lighting direction and colour temperature, colour grade reference, and a cinematic director or film that anchors the visual language. Scene duration and aspect ratio should also be specified where the tool supports them. Every unspecified element gets filled by the model's training average.
How is a dialogue scene prompt different from an action scene prompt?
Dialogue scenes require precision in character proximity, eyeline relationship, and the micro-physical performance that carries subtext — information action scenes don't need. A dialogue prompt must specify which character holds frame, where the camera axis falls relative to the eyeline between characters, and how the physical relationship between them expresses the emotional dynamic. These scene grammar decisions determine whether a dialogue scene feels tense, intimate, or alienated — and generic prompts leave all of them to chance.
Ready to try it?
Dialogue / Character Film Prompt for Claude
Three cinematically calibrated prompt variants per scene, plus a master character reference for shot continuity. Works with your existing Claude account.
Get the skill $15 · instant download · 7-day refund

Put this to work: the Dialogue Character Film Prompt skill for Claude turns everything above into one guided workflow you run in a normal Claude chat. Not ready to buy? Start with a free Claude skill and see how it works first.

Tags AI Film Sora Runway Film Prompts Claude AI AI Skills
Free skill
Try NovaKit before
you spend a dollar.

Get the LinkedIn Post Engine free — the same skill that runs live trend research before every post. Drop your email and it lands in your inbox in seconds.

💼
LinkedIn Post Engine
Social · normally $9 · free today
Live trend research before every post
Hook variants calibrated to what's converting this week
Works on a free Claude account

No spam. No account. Unsubscribe any time.