Skip to content
CodeAndBuild LogoCodeAndBuild

AI Art

How to Make a 15-Minute AI Love Story Video Content Creators Can Finish

A timed love-story structure, a cinematic image prompt, and a shot list for a 15-minute romance video you can edit from short AI clips.

CodeAndBuild Team8 min read
  • AI Art
  • Video
  • Story
  • Creators
On this page
  1. The story in one line
  2. Fifteen minutes, six chapters
  3. Prompt for the cover still
  4. Turn each chapter into clips
  5. How to finish it as a creator

A 15-minute love story is an edit, not one generated clip. Video models still work best in shots of about 5 to 10 seconds. You write the story, generate a still for each beat, turn the important beats into short clips, and cut them under one voiceover. The picture on this page is the cold open: two adults on a rainy bridge.

The story in one line

They met on a rainy bridge, built a quiet life, almost missed each other, and chose to come back to the same lamp. Keep the cast to two adults. A third face is what makes a long AI video fall apart.

Fifteen minutes, six chapters

TimeChapterWhat the viewer feels
0:00–0:45Cold openThe bridge, the lamp, the almost-touch. No names yet.
0:45–3:00MeetHow they ended up on that bridge the first time.
3:00–6:30Ordinary daysCoffee, a shared walk, one joke only they know.
6:30–9:30The crackA missed train, an unsent message, distance.
9:30–12:00The choiceEach of them alone, then one decision to return.
12:00–15:00ReturnSame bridge, same lamp, a different ending.

Prompt for the cover still

Generate this first and lock the look. Every later shot should repeat the coats, the rain, and the lamp color so the edit feels like one film.

still-prompt.txttext
Cinematic photoreal wide still, 16:9. Two adults in their late twenties, fully clothed, standing on a rain-wet city bridge at blue hour. She wears a camel coat, he wears a dark wool jacket. They face each other a step apart under a warm streetlamp, cool city bokeh behind them, wet pavement reflections, shallow depth of field, anamorphic, elegant and emotional, no text, no logos.

Turn each chapter into clips

Plan about 12 to 18 shots. A 15-minute cut does not need 15 minutes of generated footage. Hold stills, add a slow push-in, and let the voiceover carry the minutes between the moving shots.

  1. Cold open, 8 seconds. Dolly in on the two of them under the lamp. Use the still as the first frame.
  2. Meet, 8 seconds. The same bridge, earlier, one of them arriving while the other is already there.
  3. Ordinary days, 3 clips. A cafe window, a shared umbrella, hands on a railing. Same wardrobe.
  4. The crack, 2 clips. An empty seat on a train. A phone face-down on a table.
  5. The choice, 2 clips. Each person alone in their own room, then one of them putting the coat on.
  6. Return, 10 seconds. The bridge again. They stop a step apart. End before a kiss if the faces start to drift.

Clip prompt pattern

clip-prompt.txttext
8-second cinematic clip, 24 fps, 16:9. Same two adults as the reference still: camel coat and dark wool jacket, rainy bridge, warm streetlamp, blue city bokeh. One camera move only, a slow dolly forward. Natural motion, faces stay consistent with the reference, no text, no logos.

How to finish it as a creator

  1. 01

    Write the voiceover first

    About 1,900 to 2,100 words reads near 15 minutes at a calm pace. Record that before you generate clips, so the pictures follow the words.

  2. 02

    Lock one reference still

    Feed the bridge image into every clip that shows their faces. New faces in chapter four are the usual failure.

  3. 03

    Cut for the sentence, not the clip

    If a generated shot looks wrong after two seconds, use those two seconds. The voiceover is what makes the runtime.

  4. 04

    Label it

    Say in the description that the pictures are generated. Do not present the couple as real people.

More guides