BC

BityClips

Prompt pack

All prompts

5 prompts · Text-to-video generation

Text-to-video prompts for Sora, Runway, Pika and Vidu

You want generated clips that are usable as b-roll instead of uncanny five-second experiments.

Text-to-video models respond to camera language, not story language. A prompt that reads like a cinematographer's note — shot size, lens, movement, lighting, subject motion, duration — produces a usable clip far more often than a prompt that reads like a scene description. These are structured that way.

B-roll for faceless documentariesAbstract or historical shots that stock libraries do not haveEstablishing shots and transitions

The prompts

1. Cinematographer-style clip prompt

Sora, Runway, Pika, or Vidu

The default structure for a single generated shot

[SHOT SIZE] of [SUBJECT] [DOING WHAT], [CAMERA MOVEMENT], shot on [LENS/FORMAT], [LIGHTING], [TIME OF DAY], [MOOD], [COLOR PALETTE], [SECONDS]-second continuous take, no cuts, no text, no people looking at camera.

In practice: Wide aerial shot of a container ship cutting through calm water, slow forward drone push, shot on a 35mm anamorphic lens, low golden hour side light, dawn, quiet and vast, muted blue and amber palette, 5-second continuous take, no cuts.

2. Script line to shot prompt

Claude or ChatGPT, then a video model

Converting narration into generatable clips

Here are lines from my video script: [PASTE LINES]

For each line, write a text-to-video prompt using this structure: shot size, subject, subject motion, camera movement, lens, lighting, mood, palette, duration, negative constraints.

Rules: no dialogue, no on-screen text, no recognisable public figures, no logos or brands, no complex hand movements, no more than one subject per shot. If a line cannot be shown without those, say so and suggest a visual metaphor instead.

In practice: The 'cannot be shown' flag saves the credits you would burn on impossible shots.

3. Consistent look across a sequence

Claude or ChatGPT, then Runway or Sora

Multiple clips that belong to the same video

I need [N] clips for one sequence. Define a fixed style block — lens, film stock or grade, lighting logic, palette, and grain — that will be appended verbatim to every prompt.

Then write [N] prompts that vary only subject, motion, and shot size, each ending with that identical style block.

Subjects: [LIST]

In practice: Style drift between shots is the fastest way to make generated b-roll look generated.

4. Image-to-video motion prompt

Runway, Pika, or Vidu

Animating a still you already like

Animate this still image. Motion: [SPECIFIC MOTION — e.g. slow parallax push in, dust drifting through the light shaft, water rippling in the lower third]. Keep everything else static. No camera roll, no zoom out, no new objects entering frame, no morphing of the subject. [SECONDS]-second loop.

In practice: Naming what must stay still matters more than naming what should move.

5. Debugging a failed generation

Claude or ChatGPT

Fixing warped, flickering, or morphing output

This prompt produced [DESCRIBE THE FAILURE — warped hands, flickering background, subject morphing mid-shot, camera drifting]. Prompt used: [PASTE PROMPT].

Rewrite it to avoid that failure mode. Reduce the number of simultaneous motions, simplify the subject, shorten the duration if needed, and add explicit negative constraints. Explain which change targets which failure.

In practice: Most failures are too many simultaneous motions in one shot. Split it into two clips.

Variables to fill in

  • [SHOT SIZE] — wide, medium, close, macro, aerial
  • [CAMERA MOVEMENT] — push in, pull back, pan, orbit, static
  • [SECONDS] — shorter generations fail less often

What actually improves output

  • One subject, one motion, one camera move per clip. Stacking them is what produces morphing.
  • Write a reusable style block and append it verbatim to every prompt in a sequence.
  • Generate at 3-5 seconds and cut in the edit. Longer takes degrade near the end.

Mistakes that ruin the result

  • Writing story instead of camera direction — the model has no narrative memory.
  • Asking for on-screen text or legible signage. Add it in the editor.
  • Prompting hands, crowds, and fast motion together in one shot.

Tools that finish the job

Where these prompts go next

A prompt produces the script, the shot list or the markup. These are the tools that turn that output into a finished video.

Sora

High-fidelity text-to-video generation.

Limited access

View tool profile →

Runway

Creative suite for generative video, image, and editing.

Starts at $12/mo

View tool profile →

Pika

Text-to-video generation with playful motion.

Free, paid tiers available

View tool profile →

Vidu

Text-to-video generator focused on cinematic motion.

Free credits, paid tiers

View tool profile →

More prompt packs