Vertical feed delivery

AI Drama for TikTok & Reels

Updated September 20, 2026

Quick answer: three of the four things a short-form feed demands are production decisions, not export settings: the 9:16 frame is set before any art is generated, the hook is built into the episode breakdown, and the 60–120 second length is paced in the outline. Captions are the exception and are deliberately left off the video, because no current model spells reliably and burned-in text would freeze one language into the file.
Still from “The Last Envelope”, an AI short drama generated end-to-end by SceneMixer (6 characters, 1 episode)
Sample: “The Last Envelope” — Realistic HD, 1 episode, produced end-to-end by SceneMixer from the script · watch the episodes

What the platforms actually demand of the file

Short-form feeds want a tall frame, a fast opening, speech that reads without sound on, and a length that fits a single sitting. Those are production decisions, not export settings, and three of the four have to be made before any video exists.

Feed requirements and where they are decided
RequirementDecidedHow
9:16 verticalStep 1, before any artAspect ratio is a project setting; reference art and every shot are composed for it. Not a crop.
Immediate hookEpisode breakdownOpenings start on the problem; episodes end unresolved, with the kind of ending rotating.
60–120 secondsEpisode outlineShort-drama pacing caps at 240s; vertical series usually settle in the 60–120 band.
CaptionsAfter exportDeliberately not burned in — no video model spells reliably, so text goes on at the platform.
Dialogue in the viewer's languageStep 1Language is chosen once and the lines are authored in it, rather than dubbed afterwards.

Why captions are not rendered into the video

Two reasons, and neither is laziness. Current video models cannot produce reliable lettering — asked for text they return plausible-looking garble, which is why nothing in this pipeline renders words in frame at all. And even if they could, burned-in captions would freeze one language and one style into the file: changing them would mean re-rendering shots rather than editing a caption track. So you export clean video and caption it where captions belong.

The same applies to hook cards and end cards for a series episode: overlay them on the platform, change them per post, re-render nothing.

What comes out of an export

Practical sequence for a feed series

  1. Set the frame and the dialogue language first. Both propagate into reference art and every line; changing them later means regenerating art.
  2. Read the episode list before rendering anything. Each outline states how that episode ends. Fixing endings as text is free; fixing them as video is not.
  3. Render one segment. The free preview covers up to 5 seconds. Watch it on a phone, not a laptop.
  4. Render one full episode before committing a season. At 10 credits a second, a 90-second episode is about 990 credits including assembly.
  5. Caption and publish per platform. Same clean master, different caption tracks and covers.

What it will not do

What it costs

Video from 10 credits per second of output (≈ $0.06/s at the lowest credit price), assembly 1 credit per second, shot lists 1 credit per hundred characters, analysis 4 credits per thousand. New accounts get 50 credits, 15 free reference images and one free preview. The cost breakdown prices out a full eight-episode series; the vertical page covers what changes in a tall frame.

Languages

Native dialogue in 15 languages

Every series below was produced by the same pipeline with the cast speaking that language natively: the interface, the working documents and the spoken lines are all in one language, and nothing is dubbed. Open one to hear it.

See all 15 languages

Frequently asked questions

Do I get captions burned into the video?

No, deliberately. No current video model produces reliable lettering, and burned-in text would lock one language into the file. You export clean video and add a caption track on the platform, where you can change it per post without re-rendering.

Can I turn a landscape episode into a vertical one?

Not by cropping — there is no reframing pass. Vertical output comes from setting the project to 9:16 before reference art is generated, so every shot is composed for the tall frame.

How long should a feed episode be?

Sixty to a hundred and twenty seconds is the working band for vertical short drama, and the outline paces for it. The format's ceiling here is 240 seconds per episode.

Is there an AI watermark on the output?

A small corner badge is applied at delivery by default. Paid members can switch it off for subsequent renders. Whatever a given platform requires in the way of AI disclosure is the platform's rule and worth reading yourself.

Can it upload to TikTok or Reels directly?

No. There is no posting integration — you download the finished episode and upload it yourself, which is also what lets you caption the same master differently per platform.

One clean master, captioned per platform

Set the frame, render five free seconds, watch it on your phone.

Try SceneMixer free

Credit packages from $1.49 · Pro $7.99/mo · No credit card to start

Pricing·Novel to Video·Guides·Samples·Support·Privacy·Terms·Legal·Contact·© 2026 SceneMixer