🎬 Reel Maker β€” Team Guide

Turn a short written story + your photos and clips into a finished, captioned vertical video β€” no editing skills needed. This guide explains every button and field in plain language, with the exact words you can type.

πŸ› Doing a brand deal or ad? Use the Brand / Ad mode link at the top. It works the same way, plus: you fill a short brand brief (or paste the brand's brief), upload the logo and real product photos (the tool never fakes a product or logo), tick πŸ› Product beat on the product shots, and choose the voice-over and music (AI draft or the brand's own files). It won't let you render until you've added the logo, a call-to-action, and a product shot β€” and it auto-adds an end-card + a "Paid partnership" line. You type all the ad wording and claims yourself β€” the AI never writes claims.
Where to open the tool: https://creative.kevat.ai
Works in any browser. Nothing to install. If a teammate is mid-render, just wait a moment and refresh.

What's inside

1 Β· What it does (in one minute) 2 Β· The 8 steps, start to finish 3 Β· Step 1 β€” Subject (who the story is about) 4 Β· Step 2 β€” Write the story (with template) 5 Β· The Frame Card β€” every field explained 5a Β· Picture source (Auto / Upload / AI) 5b Β· Director Note β€” the most important box 5c Β· Camera Motion β€” full keyword list 5d Β· Image Edit β€” change a photo with words 5e Β· Lip Sync β€” make a face talk 5f Β· Voice-Over β€” narration over the video 5g Β· Duration & Video Start 5h Β· Picture & Video "model" pickers 6 Β· Style, Quality, Mood, Captions 7 Β· Music 8 Β· Cost panel β€” read before you press Generate 8b Β· Iterate without re-rendering (redo Β· approve Β· live clips) 9 Β· How to build a story that lands 10 Β· One-page cheat sheet (all keywords) 11 Β· A full worked example 12 Β· Quick fixes (FAQ) 13 Β· Brand / Ad Mode β€” full walkthrough 13a Β· Brand brief panel 13b Β· Brand kit & assets 13c Β· Audio β€” VO & music 13d Β· Writing the ad script & marking product beats 13e Β· Mandatories hard-block 13f Β· Brand mode quick fixes 14 Β· Studio Mode β€” type a brief, get a reel 14a Β· Identity library (Talent & Product) 14b Β· Brief β†’ shots 14c Β· Per-shot controls

1 Β· What it does (in one minute)

You give it a story written as short lines and a folder of photos/videos. For each line it decides what should be on screen, how the camera moves, and the mood β€” then builds one 9:16 vertical video with captions and music, ready for Reels / Shorts.

Think of it as an automated video editor with a film director's brain. Each line of your story becomes one shot. A 10-line story becomes a 10-shot reel.

2 Β· The 8 steps, start to finish

1
Subject β€” type the person's name + a short look description.
2
Start your reel either by pasting a frame-by-frame script, or by choosing I have a story and pasting raw story text/notes/transcript. The AI draft path fills an editable script first; review it before preview/render. Then click πŸ“ Browse folder… and choose the folder with your photos/videos (they upload to the tool). Manual scripts use Parse Frames β†’; AI drafts parse after the draft is created.
3
Frame cards appear β€” one per line. Tweak the picture, mood note, and camera on each.
4
Target length β€” pick 30/45/60/90s or Auto.
5
Style & Quality β€” start in Dev (cheap) while testing.
6
Music β€” none, upload, or auto-generate.
7
Check the πŸ’° cost panel.
8
Generate β†’ watch the log β†’ Download .mp4.

3 Β· Step 1 β€” Subject (optional)

You can skip this entirely. Leave both fields blank and the tool works out who is on screen β€” their age, gender and look β€” from the story you write. There is no fixed "default person"; the story decides.
FieldWhat to typeExample
Name (optional)The main person's first name, if you want to pin it.Lalita
Description (optional)A short look so AI-made pictures stay consistent (helps most with the Consistent face tick). Fine to leave blank.woman, 30s, strong features, traditional clothing
More than one person in the story? When a line quotes someone β€” e.g. "My son asked, 'Mom, where is father gone?'" β€” the tool spots it and shows that person (the kid's face, the right age) and, if lip-sync or voice-over is on, their voice β€” not the narrator's. Each frame card gets a 🎭 Speaker dropdown to fix any line, and a 🎭 Cast voices panel lets you pick a voice per person.

4 Β· Step 2 β€” Write the story

The story is plain text. Start each shot with the word Frame and a number, then write the line that should appear on screen. That's it.

Copy-paste template

Reels

Frame 1
From a farmer to a model… the story of a desi girl with big dreams.

Frame 2
At 19, I got married into a traditional Rajasthani home.

Frame 3
A kisan ki beti β€” I rode tractors even after marriage.

Caption:
Write your full Instagram caption here. (This does NOT appear in the video.)
RulePlain meaning
One Frame = one shotKeep each line short β€” it shows on screen as a caption while the camera moves.
Numbers don't need to be in orderThe tool reads top-to-bottom, not the numbers.
The Caption: block at the bottomIs your Instagram post text. It is never shown in the video β€” write the long version here.
Keep on-screen lines under ~15 wordsThe camera moves before longer text can be read.
You don't have to type anything fancy. Everything below (picture, mood, camera, lip-sync) can be set by clicking on the frame card after you parse. The bracket tags shown later are just the typed shortcut for the same buttons.

Adding your own photos & videos

In story mode, start with either I already have a frame script for full control, or I have a story to let the AI create an editable frame draft from raw notes/transcript. The AI does not render directly from the raw story; it fills the script/frame cards so you can review first.

Click πŸ“ Browse folder… (next to the start panel) and pick the folder on your computer that holds the photos/videos for this story. The tool uploads them and matches the right one to each shot automatically.

Tick πŸ€– Smart-match after uploading and the tool reads what's in each photo (people, places, even names/text) and places the best photo on each line β€” instead of just going in order.
Tick 🎞 Multi-shot coverage to make longer beats richer: instead of one photo for the whole line, the tool covers it with the main shot plus 1–2 supporting B-roll stills (it picks ones that also fit that line), played as quick sub-shots with gentle camera motion under the same caption.

So a 9-second beat might become three 3-second clips (main photo + two B-roll stills), making the video feel more edited and less like a slideshow.

Note: it only kicks in when a beat is long enough and you have spare matching photos. It uses extra credits (~1.5–2Γ— on the beats it covers), so turn it on when you want that polished, multi-cut feel for important story moments.

5 Β· The Frame Card β€” every field explained

After Parse Frames β†’, each line becomes a card. Top to bottom you choose: the picture, then optionally a mood note, a photo edit, a camera move, a lip-sync toggle, and the length.

πŸ’‘ Suggestions β€” pick, don't type. Under the mood note and photo edit boxes you'll see small clickable suggestions; click one to drop it in (then edit), or ignore it. Camera is a dropdown pre-set to the AI's best pick β€” just change it if you want something else. And after πŸ‘ Preview Stills, each frame gets a ✨ Suggest from image button: the tool looks at the actual generated frame and sets the best camera move + offers a director note tailored to what's really on screen. You choose to keep or change it. (Photo edits stay yours β€” the tool doesn't auto-suggest those.)

5a Β· Picture source β€” where the shot's image comes from

ButtonUse it when…
AutoYou're happy with the photo it auto-picked from your folder.
πŸ“ From FolderUse the matched file from your folder (shown as a green badge).
πŸ“· Upload PhotoPick any photo or video from your computer for this shot.
🎨 AI PortraitNo good photo? The AI creates a person/portrait for this moment, matching the Subject + your note.
πŸ–Ό AI SymbolicThe AI creates objects/scenery with NO people β€” perfect for sad/abstract beats (illness, loss, a turning point) where showing a face feels wrong.
Golden rule: real photos and clips always look best β€” use them whenever you have one. Use AI Portrait when you have no photo, and AI Symbolic for pain or abstract moments (show the empty room, not the person crying).

5b Β· Director Note β€” the most important box

This tells the AI director the feeling and look of the shot. Write it like you're briefing a photographer: what's in frame + how they feel + the light.

Feeling you wantWhat to type in the note
Pride / TriumphHead high, direct gaze, warm gold backlight, full confidence
Grief / LossNo eye contact, slumped posture, cold blue-grey light, objects rather than face
LongingEyes looking just off-frame, window light, half-turned away
DeterminationJaw set, hands busy, warm practical light, grounded posture
Fear / UncertaintyShadow across the face, shallow focus, background looms out of focus
InnocenceYoung face, soft diffused light, looking up slightly, clean background
Turning point / HopeSpark in the eyes, phone glow in a dark room, a small smile starting
Vague noteSpecific + emotional note
show her being sadMedicine bottles on a windowsill, no person, cold winter light, the silence of an empty room
the ramp walkHead high, back straight, heels clicking, full confidence, warm gold backlight β€” she earned this

5c Β· Camera Motion β€” pick from the dropdown

Camera is a dropdown, not a text box β€” no typing. It's pre-set to the AI's best pick for that beat (you'll see a ✨ auto note). To change it, just choose another move from the list below. Leave it on ✨ Auto to let the director decide at render. After πŸ‘ Preview Stills, the ✨ Suggest from image button re-picks the best move by looking at the actual frame. Here's what each move does:

Type thisWhat you'll seeBest for
360 orbitCamera circles all the way aroundHero reveal, triumph
bullet timeFreeze + orbit (Matrix style)The single peak moment
crash zoom inFast dramatic zoom toward the subjectShock, surprise, realization
crash zoom outFast zoom awaySudden scale, "the world opens up"
dolly in / push inSmooth glide toward the subjectIntimacy, building emotion
dolly out / pull backSmooth glide away, revealing surroundingsLoneliness, scale, isolation
crane upCamera lifts upwardVictory, freedom, rising
crane downCamera lowersWeight, defeat, gravity
tilt up / tilt downCamera angles up or downLooking to the sky / looking down
overheadBird's-eye view from directly aboveVulnerability, isolation
dutch angleTilted, off-kilter horizonTension, unease, something's wrong
Hitchcock zoomVertigo effect (zoom + pull)Dread, disorientation
extreme close on eyesMacro on the face/eyesDeep emotion, connection
arc left / arc rightCamera sweeps in a partial circleGentle reveal, momentum
handheldSlightly shaky, human feelRaw truth, documentary realism
staticNo movement at allStillness, weight, gravity
super 8mmVintage film-grain lookMemory, flashback, nostalgia
whip panFast blurred swipeEnergy, time passing, scene change
One move per shot. Keep it to a single gentle action (e.g. slow push in). Over-describing motion is the #1 cause of weird, melty faces. And remember: a shot is either a camera move or a talking face (lip sync) β€” not both.

5d Β· Image Edit β€” change a photo with words

The ✏️ Image edit box changes the picture before it animates. Type a plain instruction.

Type thisResult
add thunderstorm and dark cloudsStorm added to the sky
make the lighting warmer and goldenWarm sunset tone over the whole image
add rain on the windowRain streaks added
add soft morning fogAtmospheric fog layer
cold grey winter light and frostTurns a scene cold and sombre

Big mood changes ("add storm", "warmer light") work great. Precise moves ("shift her to the left") are less reliable β€” that's expected.

5e Β· Lip Sync β€” make the subject speak

Tick πŸŽ™ Lip Sync and that shot becomes a talking face that says the line aloud in a real voice. A voice menu appears β€” pick one or leave the default.

Use it on 2–4 shots max β€” the spoken turning point and the closing line. Don't lip-sync everything; mix talking beats with silent cinematic beats.

5f Β· Voice-Over β€” narration over the full video

Check β˜‘ Voice-Over Track at the top to make the entire video narrated. Each caption gets read aloud in a chosen voice, creating a full voice-over narration layered under the music.

Lip Sync vs Voice-Over: Use Lip Sync (β˜‘ on a frame) when you want a character speaking on screen, moving their lips. Use Voice-Over Track when you want a narrator telling the story over cinematic shots β€” think documentary or portfolio reel.

5g Β· Duration & Video Start

FieldMeaning
DurationHow long the shot lasts. Leave blank and it's auto-calculated from word count (longer line = longer shot). Lip-sync shots set this themselves.
Video start (only for video sources)Skip the first few seconds of a clip. e.g. 3 starts the clip 3 seconds in β€” handy when the good part is mid-way.

5h Β· Picture & Video "model" pickers

"Model" = which AI engine makes the picture or the motion. Leave both on Auto and the tool picks the best one for each shot, cheap while testing and premium for the final. Only override if you want a specific look.

Picture models

NamePlain meaningTier
SeedreamCheap, fast pictures β€” great for testingdraft
Nano BananaPremium, very photo-real facespremium
FluxPremium portraits & objectspremium
GPT ImageBest for objects, symbolic scenes, and any text-in-imagestandard

Video (motion) models

NamePlain meaningTier
Kling StandardSolid animation, great value β€” the testing defaultdraft
Kling ProHighest-quality Kling motionpremium
HiggsfieldMost cinematic camera-move presetspremium
SeedanceCinematic, multi-shot feelpremium
VeoBest built-in audio/dialoguepremium
HailuoWide landscapes, realistic motionpremium
Ken BurnsSimple zoom, free, no AI β€” for free testsfree
99% of the time: leave both on Auto. Auto already uses cheap engines in Dev and premium in Production, and never re-generates your real photos.

6 Β· Style, Quality, Mood & Captions

Quality tier β€” "test cheap, finish expensive"

TierUse when
Dev cheap, 5s shotsAlways start here. Cheap draft to check your story, captions, and timing.
Production premium, up to 9sOnly for the final render once everything looks right.

Mood / colour palette

MoodLook
Warm NostalgicAmber, golden-hour, slightly vintage
Cold StruggleBlue-grey, overcast, deep shadows
TriumphantRich golds & saffron, bright, high saturation
DefaultLet the AI choose per shot

Orientation, transition & captions

SettingOptionsPick
OrientationPortrait 1080Γ—1920 / Landscape 1920Γ—1080Portrait for Reels/Shorts/TikTok
TransitionCrossfade / Hard CutCrossfade for smooth, Hard Cut for punchy
Caption fontBaskerville / Montserrat / Satoshi* / Arial / Georgia / HelveticaBaskerville for storytelling, Montserrat for modern & clean
Caption size24–96 pt52 default; 60–70 for dramatic
Caption colourWhite / Yellow / BlackWhite β€” always readable
Caption position (default)Bottom / Middle / TopBottom for Reels β€” override per frame if a shot needs it elsewhere
Burn captionsOn / OffUntick for a clean video with no subtitles. Voice-over (if on) still plays.
Max lines per caption (default)No limit / 1 / 2 / 3Caps how many lines a caption takes; long text auto-shrinks to fit. 1–2 lines keeps it clean.
*Satoshi is a paid font β€” it only works once a licensed font file has been added to the app (ask your developer). Montserrat works out of the box.
Per-frame caption control. The position and max-lines pickers in Style & Quality are the defaults for the whole reel. On any frame card you'll see a πŸ’¬ Caption row with its own Position and Lines dropdowns β€” set them to move just that one caption (e.g. push it to the Top when the subject's face is at the bottom of the shot, or force a punchy line to 1 line). Leave them on "default" to follow the global setting. All of this works the same in Brand / Ad mode.

7 Β· Music

OptionWhat it does
No MusicSilent video β€” add your own later.
Upload MusicUse your own MP3/M4A/WAV. It loops and fades out at the end.
Auto-GenerateDescribe a mood and it makes a track (~2–3 min). Wait for βœ“ Music ready before Generate.

Good auto-music prompts: Emotional Bollywood instrumental, struggle to triumph, sitar and tabla, no lyrics Β· Melancholic Rajasthani folk, raw acoustic

Combine music + voice-over: Enable Voice-Over Track (above) to layer narration on top of your music. Music plays at 25% volume, so the voice is clear and the music sits underneath.

8 Β· The πŸ’° cost panel β€” read it before Generate

After you parse, a cost estimate appears and updates live as you change settings. Check it before pressing Generate. Want it cheaper? Use Dev tier, prefer real photos/videos, and keep lip-sync to a few shots.

πŸ’³ AI Credits panel (top of the page). Shows how much credit is left on each AI service so you can recharge before a render fails mid-way. Live balances show today for ElevenLabs (voice), Kling (video), Suno (music), and fal.ai (images/video). A red "recharge!" means that wallet is empty. Some services (OpenAI, Gemini, Higgsfield, Hedra, SyncLabs) don't publish a balance β€” they show "β€”" and you check their own dashboard. Hit ↻ Refresh after topping up.

8b Β· Iterate without re-rendering β€” redo, approve, live clips

You don't have to render the whole video and hope. Treat the tool as a draft machine you steer: preview the stills, fix the frames you don't like, approve the rest, and only pay to animate what you've approved. Three controls make this fast.

The recommended 3-pass workflow
1. πŸ‘ Preview Stills (free) β€” see every image.
2. Fix & approve per frame β€” πŸ”„ redo any image you don't like; untick frames you're not happy with yet.
3. β–Ά Generate in Dev β€” only approved frames get paid animation; you watch each clip appear live.

πŸ”„ Redo still β€” fix one image, not the whole video

After Preview Stills, each frame card shows a πŸ”„ Redo still button. Change that frame's note, photo, image edit, or camera β€” then click πŸ”„. Only that one image regenerates (a few seconds), and it swaps into the card. Nothing else is touched, and you don't pay to re-animate anything.

βœ“ Approve for animation β€” only pay for the frames you like

Each frame card has a βœ“ Approved for animation tick (on by default). Untick any frame you're not happy with. When you press Generate:

This is the cheap way to test direction: approve only 2–3 hero frames first, render, see how they feel, then come back and approve more. You're never forced to pay for all frames at once.

🎬 Live clip reveal β€” watch it build

While a render runs, each frame's clip appears in its card the moment it's done β€” you don't wait for the whole video. The card shows ⏳ animating…, then flips to a looping 🎬 clip ready preview. If an early clip looks wrong you can stop, fix that frame with πŸ”„, untick it, and re-run β€” without having waited for all ten.

Timeline, safe-zone, posting kit, and export

9 Β· How to build a story that lands

ShotJob
Frame 1Hook β€” who is this, why watch? Use your best clip or a crash-zoom.
2–3Context β€” the before.
4–5Conflict β€” the problem / loss (use AI Symbolic for pain).
6–7Lowest point, then the spark (great lip-sync moment).
8–9Turning point β†’ triumph (real footage of the win; save 360 orbit for here).
10Resolution β€” strong, present-day, direct gaze (great closing lip-sync line).

10 Β· One-page cheat sheet (print this)

Picture source: Auto Β· From Folder Β· Upload Β· AI Portrait Β· AI Symbolic

Camera moves: 360 orbit Β· bullet time Β· crash zoom in/out Β· dolly in (push in) Β· dolly out (pull back) Β· crane up Β· crane down Β· tilt up/down Β· overhead Β· dutch angle Β· Hitchcock zoom Β· extreme close on eyes Β· arc left/right Β· handheld Β· static Β· super 8mm Β· whip pan

Moods: Warm Nostalgic Β· Cold Struggle Β· Triumphant Β· Default

Picture models: Seedream (cheap) Β· Nano Banana (photo-real) Β· Flux Β· GPT Image (objects/text)
Video models: Kling Standard (default) Β· Kling Pro Β· Higgsfield Β· Seedance Β· Veo Β· Hailuo Β· Ken Burns (free)

Quality: Dev (test cheap) β†’ Production (final)
Features: Lip sync (2–4 shots) Β· Voice-Over track (full narration) Β· Multi-shot coverage (longer beats split into B-roll)
Iterate: πŸ”„ Redo still (fix one image) Β· βœ“ Approve per frame (pay only for approved; rest = free Ken Burns) Β· live clip reveal while rendering
Edit examples: add thunderstorm Β· warmer golden light Β· add rain on the window Β· morning fog

Optional: typing it in the story instead of clicking

Every button has a typed shortcut you can put under a frame line. The buttons do the same thing β€” use whichever you prefer.

Type under the lineSame as the button…
[photo: filename.jpg]Use a specific file
[photo: ai_portrait] / [photo: ai_symbolic]AI Portrait / AI Symbolic
[note: ...]Director Note
[camera: 360 orbit]Camera motion
[edit: add storm]Image Edit
[lipsync: yes]Lip Sync toggle
[duration: 8] Β· [start: 3]Duration Β· Video start

11 Β· A full worked example

A real story turned into a 10-shot reel. Paste it, set the folder, hit Parse β€” then read why each choice was made.

Reels

Frame 1
From a farmer to a model… the story of a desi girl with big dreams.
[photo: ai_portrait]
[note: Strong proud Rajasthani woman, direct gaze, chin up. Golden hour, dust in air. A HERO, not a victim.]
[camera: crash zoom in]

Frame 5
After COVID, I was diagnosed with arthritis. I lost my hair, my confidence… and myself.
[photo: ai_symbolic]
[note: NO person. Medicine bottles on a windowsill, a hairbrush with fallen strands, cold blue-grey light. The objects carry the grief.]
[edit: cold grey winter light and frost on the window]
[camera: static]

Frame 7
Bedridden, I watched modelling videos thinking, "one day I'll walk the ramp."
[photo: lalita_face.jpg]
[lipsync: yes]
[note: The turning point β€” she says it herself. Phone glow on her face in a dark room. A small smile starting.]

Frame 8
I fought back β€” and the girl who couldn't stand walked the ramp in heels.
[photo: ramp_walk.mov]
[start: 3]
[note: Real ramp-walk clip. Head high, full confidence. The peak. Skip the first 3 seconds.]
[duration: 8]

Frame 9
I won Mrs. Rajasthan 1st runner-up.
[photo: crown.jpg]
[note: Pure victory β€” crown, sash, stage. She earned this. High saturation, saffron and gold.]
[camera: 360 orbit]

Caption:
From the farm to the ramp β€” the full story for your followers goes here.
FrameWhy
1No strong opening photo β†’ AI Portrait hero shot. Crash zoom stops the scroll in the first 1.5s.
5Illness shown as objects, not a face (AI Symbolic) + cold edit + static camera = the weight of stillness.
7Lip sync on the spark β€” hearing her say it lands harder than text.
8Real video of the win beats any AI. start: 3 skips the boring intro.
9360 orbit saved for the single highest moment.

12 Β· Quick fixes (FAQ)

ProblemFix
Face looks melty/distortedSimplify the camera move (one gentle action), or use a real photo.
Wrong photo on a beatClick πŸ“ Upload or From Folder on that card and pick the right one.
Caption too long / gets cutShorten the line to under 15 words.
Costs look highSwitch to Dev, use real photos, cut lip-sync to 2–4 shots.
Lip-sync shot came out silent/animatedThat shot fell back safely β€” check the line isn't empty and try again.
Testing for freeSet Video model to Ken Burns + Quality to Dev β€” zero AI cost.
Don't like one imageClick πŸ”„ Redo still on that frame after Preview β€” only that image regenerates.
Only want to pay for some framesUntick βœ“ Approved on the frames you're unsure about β€” they animate free (Ken Burns); the cost panel drops.
Don't want to wait for the whole renderClips appear in each card as they finish β€” watch the early ones, stop & fix if a frame looks wrong.
Wrong speaker on a quoted lineUse the 🎭 Speaker dropdown on that frame card to pick the right person.
AI chose wrong gender/age for a characterFill the Subject β†’ Description field with gender + age, or leave blank and the AI infers from the story.

13 Β· Brand / Ad Mode β€” full walkthrough

Click πŸ› Brand / Ad mode in the header (or go to /brand) to switch. The engine is identical β€” same pipeline, same frame cards, same models β€” but the UI gains a brand brief panel, real-asset uploads, mandatory checks, and a product beat toggle per shot.

Iron rule: you type every word of ad copy yourself. The AI never writes, rephrases, or invents marketing claims, taglines, or on-screen text. It only makes visuals. Paste the brand's exact words into the fields below.
1
Paste the brand brief (optional) β†’ click Extract fields β†’. Fields fill in with the brand's exact words β€” never paraphrased. Only empty fields are filled; anything you've already typed is left alone.
2
Upload the logo (required) and optionally product photos/videos.
3
Paste the ad script in the beats box, then Parse Frames β†’. Each beat = one frame card. Tick πŸ› Product beat on the shot(s) where the real product/logo appears.
4
Choose audio β€” AI announcer (draft) or brand-supplied VO file; AI music or brand-supplied cleared music.
5
πŸ‘ Preview Stills first (free β€” no animation) to check the look and product-beat frames.
6
β–Ά Generate Ad β€” the tool checks all mandatories before spending any credits and shows a checklist if anything is missing.

13a Β· Brand brief panel

Option A: Paste the full brief text β†’ Extract fields β†’ (parses verbatim, fills empty boxes only).
Option B: Type directly into the fields. Either way, what you see in the boxes is what gets used.
FieldWhat to put hereNotes
Brand nameThe brand's nameAppears in the "Paid partnership" disclosure line
ProductThe product category or nameUsed in visual context β€” not on screen
ObjectiveWhat the ad should achievee.g. product awareness, trial drive
Key messageThe brand's exact claim β€” verbatimNever rephrased by AI
CTA text (required)The call-to-action line shown on the end carde.g. "Try it now Β· Link in bio"
CTA linkURL or "Link in bio"Shown on the end card
TaglineBrand tagline (optional)Optional β€” only set if the brand supplied it
Brand colourHex code for the end carde.g. #0a7d33
Announcer scriptExact voice-over linesRead aloud as-is β€” never rewritten

13b Β· Brand kit & assets (real-only β€” never AI-generated)

AssetRequired?What it does
Logoβœ… RequiredPlaced on the end card; optionally shown as a logo bug in the corner throughout the video
Product photos/videosRecommendedUsed on frames you mark πŸ› Product beat β€” never AI-replaced
Show logo in cornerOptionalTick to burn a small logo bug into every shot; pick corner (top-right default)
Product beat frames are real-only. When you tick πŸ› on a frame, it uses your uploaded product photo/video exactly as-is β€” the AI never generates a product shot. If you haven't uploaded one, that frame gets a plain visual instead.

13c Β· Audio β€” VO & music

Brand mode separates the announcer voice-over from the background music. Both are decided per-project.

SettingOptions
Announcer VOAI draft β€” reads your announcer script in a chosen ElevenLabs voice (fast, for review) Β· Brand-supplied audio β€” upload the final signed-off VO file
Background musicAI music β€” auto-composed from the ad mood Β· Brand-supplied music β€” upload their cleared track
VO plays over music. The announcer is at full volume; music ducks to 18% underneath automatically β€” no ffmpeg knowledge needed.

Need an ElevenLabs voice ID? Paste it in the VO voice field. Leave blank for the default voice.

13d Β· Writing the ad script & marking product beats

Write the script exactly like a story script β€” one caption per Frame. The caption is the on-screen text the brand supplies. After parsing, each frame card gets a πŸ› Product beat toggle.

Beat typeWhat to do
Lifestyle / story beatWrite the brand-supplied line. Leave product beat off. The AI makes the visual.
Product / logo shotTick πŸ› Product beat. Upload a real product photo in the Brand kit. That photo is used as-is β€” no AI generation.
End-card beatAuto-generated by the tool from your CTA text, logo, and brand colour. You don't write it β€” it's appended automatically.
Best structure for a 15–30s ad:
Frame 1: Lifestyle hook (relatable scene, brand-neutral)
Frame 2–3: Problem or desire
Frame 4: Product introduction (πŸ› Product beat β€” your product photo)
Frame 5: Key message (brand's exact words on screen)
End-card: Auto-generated CTA

13e Β· Mandatories hard-block (the checklist)

If you click Generate Ad with something missing, the tool refuses to spend any credits and shows a red checklist. You cannot generate until all boxes are ticked. This protects against accidentally publishing an ad without a logo or disclosure.

MandatoryHow to clear it
Brand logo (upload the logo file)Upload a logo file in the Brand kit section
Call-to-action textFill the CTA text field in the Brand brief panel
At least one πŸ› Product beatTick the πŸ› Product beat toggle on at least one frame card
Disclosure (auto-added)This one is always satisfied β€” the tool adds "Paid partnership with [Brand]" automatically
The Instagram "Paid Partnership" label is separate. The tool burns a text disclosure into the video file. You still need to set Instagram's native Paid Partnership with Brand label when posting β€” that's done inside the Instagram app, not here.
Consent and rights are also mandatory. Brand mode includes a confirmation that you have consent / likeness / content rights for real people, product assets, logo, VO, and music. The render is blocked until that is checked.

13f Β· Brand mode quick fixes

ProblemFix
Hard-block checklist appears when I click GenerateComplete each item in the red checklist β€” upload the logo, fill CTA text, tick a product beat.
Product photo isn't showing on the product beat frameUpload a product photo in the Brand kit section, then tick πŸ› on the right frame.
Extract fields filled in the wrong thingJust edit the field directly β€” extracted values are pre-fills, not locked. You always win over the AI.
VO is the wrong voicePaste an ElevenLabs voice ID in the VO voice field, or switch to brand-supplied audio.
I want to test the look without spending VO/music creditsUse πŸ‘ Preview Stills β€” it generates images only, no animation, no audio spend.
End-card text is wrongEdit the CTA text field and the brand name field β€” the end card is built from those.

14 Β· Studio Mode β€” type a brief, get a reel

Open 🎬 Studio mode from the top link. Instead of writing a script, you describe the reel in plain words and the AI drafts the shots. It uses the same engine (Preview, cost, Generate, export) plus a reusable identity library.

14a Β· Identity library (Talent & Product)

Save a Talent (a face) and/or a Product (a photo + specs) once. Studio locks them across every shot so the same person and the same product appear consistently β€” and they're reusable in future reels. On shots you mark πŸ› Product beat, the real product image is used directly (never re-generated), so logos and fine detail stay exact.

14b Β· Brief β†’ shots

Pick a scope β€” Commerce (product / fashion / jewelry ads) or General (any idea) β€” write your brief, and click ✨ Plan shots. The AI returns editable shot cards (on-screen line, camera move, shot size). Edit any card before Preview/Render. The AI may draft the on-screen lines here (you edit them); all text still passes the safety check.

14c Β· Per-shot controls

Each shot card adds a Talent/Product selector, a Negative prompt (what to avoid), and a Continuity lock (outfit/styling that must not change). Sensible defaults are prefilled; tweak per shot as needed.


Story mode workflow: Write short lines β†’ Parse β†’ set Picture + Note + Camera per shot β†’ Dev test β†’ Production final β†’ Download.
Brand mode workflow: Fill brief β†’ upload logo + product β†’ write ad script β†’ Parse β†’ tick πŸ› product beats β†’ Preview Stills β†’ Generate Ad β†’ Download.
Studio mode workflow: Save Talent/Product β†’ pick scope β†’ write brief β†’ ✨ Plan shots β†’ edit cards β†’ Preview Stills β†’ Generate Reel β†’ Download.
When unsure, leave models on Auto and start in Dev.