One character, every episode
Saved characters with pictures, motion clips and a voice; the same face and style across projects.
A complete video from scratch: script, AI presenter, B-roll, voice — no footage needed.
AI Video Creator is Film To Story without a film. You save a character once (an AI picture or a photo), choose how much of the video should be the presenter on camera, and your AI writes a script that alternates presenter segments (PR01, PR02…) with B-roll pictures and short AI video clips. The presenter is animated on your PC with free models: motion from a recorded clip or a Grok Imagine clip, lips synced to the voiceover, an AI-placed camera that punches in on key words. Optional source videos can still be added and cut like in the other tools.
The Film To Story pages, minus the film: a Character step instead of Import, a Script page with the presenter segments, and a B-roll card that also makes video clips.
Pick a saved character or make one: a picture generated by Gemini / ChatGPT / Grok from a description (with reference pictures attached), or your own photo. Add motion clips — record yourself for 10 seconds or let Grok Imagine make a clip — and tag them (calm talk, hands, lean in…).
Topic, tone, duration and language as in Film To Story, plus the presenter share (how much of the video is on camera, e.g. 50 %) and which B-roll kinds the AI may use: AI pictures, AI video clips, real footage.
The AI writes the script and a presenter list: segments with motion, camera events (zoom on a word, anchored on the face or hands) and durations that respect your share target; B-roll pictures and clips fill the rest.
Import checks the schema, the share against your target and the motion variety. Each segment shows its beat, voice and motion; Render this segment animates it right away (Grok clip motion, LivePortrait, or still picture with lips), Try picture + motion previews a different combination.
Pictures: generate with Gemini / ChatGPT / Grok or import. AI video clips: Grok Imagine makes each one from its prompt (6 or 10 s, cut to the beat at render), one by one or all at once; import your own clip instead. Real footage: the script asks, you import.
Voice engine and voice per project or from the character; per-beat regeneration; the strip shows presenter segments, pictures and clips in order; music and ducking.
The Render page shows which presenter segments are rendered and fresh; Render N segments makes the missing ones in one go. The video is cut without a film: presenter segments where they fall in each beat (lips on the words), clips trimmed exactly, pictures animated, captions word by word.
Per-format cards, the unified row (Render again · Download · Copy link · cloud state · Publish), thumbnail composer and the publish modal.
Everything runs on your PC with free models — the only accounts you need are the AI chats you already use.
Saved characters with pictures, motion clips and a voice; the same face and style across projects.
LivePortrait motion + MuseTalk lip-sync + MediaPipe, bundled (≈ 4.5 GB), no API, no subscription.
Tell the AI 30 % or 70 % on camera; the prompt computes the quota and the validator rejects scripts that overshoot.
Grok Imagine text→video from the script's prompt, right on the B-roll card, in the row's aspect.
Gemini, ChatGPT or Grok, with reference pictures attached, in one visual style.
Script-placed zoom events anchored on the face or hands, eased at render.
Script and voice in your language; captions word by word; Arabic RTL handled.
Pictures and clips are made in the background with your connected accounts, several at once.
Characters and projects restore on another PC.
Nothing is hidden behind a paywall or an "advanced" plan — these are the knobs in the app.
| Setting | Options | Notes |
|---|---|---|
| Character | Saved · picture (AI / photo) · motion clips · default voice | Reusable across projects. |
| Presenter share | 10–90 % | Target the AI must respect; validator tolerance +25 %. |
| B-roll kinds | AI pictures · AI video clips · real footage | Which the script may use. |
| Segment | Motion · camera events · duration · engine (auto / Grok clip / LivePortrait / still) | Per segment on the Script page. |
| AI clips | Grok Imagine 6 / 10 s · aspect from the row · start at | Cut exactly to the beat at render. |
| Voice · music · captions · thumbnail | As in Film To Story | |
| Format | Reel 9:16 · YouTube 16:9 | Each rendered separately. |
What the tool does not do (yet), so you are not surprised.
Free download for Windows. A topic and a picture tonight, a presenter video tomorrow.
Free · No watermark · No clip limits · Runs on your PC