| Moment selection |
Transcript, waveform, duplicate, and timing evidence; limited durable visual-performance data. |
Picks gaze, credibility, gesture peak, microexpression, product visibility, and the revealing half-second after the line. |
Performance atlas per phrase/take: emotion, eye contact, gesture, blink, energy, subject/product, technical confidence. |
| Hook |
Strong written doctrine, but no required uploaded-footage hook lab or first-frame acceptance gate. |
Picture, voice, text, and sound converge on one promise in the first event—not a title-card throat clear. |
Three real hook assemblies: result-first, contrarian/confession, sensory demo; contact sheet + 0–2.5s preview. |
| Story compression |
An agent can build beats; no required promise → proof → payoff ledger or alternate assembly. |
Moves proof before explanation, deletes correct-but-unnecessary lines, preserves only setup, escalation, objection, payoff. |
Edit-decision map with each beat’s intent, claim, proof asset, payoff, transition reason, and confidence. |
| Pacing |
Docs reject metronomic cutting; execution remains prompt-led and has no retention feedback. |
Pattern interrupts at narrative turns, holds when comprehension or emotion needs time, accelerates only for tension/payoff. |
Intentional cadence: beat role + shot-length rationale + viewer-response data linked to the timeline. |
| Sound |
Deterministic music and loudness tooling; limited productized dialogue repair, ambience continuity, or first-class split edits. |
J-cuts the next idea, L-cuts voice over proof, smooths room tone, preserves breath and tactile click/spray/keyboard as evidence. |
Five-part sound scene: dialogue, ambience, native product sound, music, sparse SFX—with independent A/V in/out. |
| Captions |
Real word timing and art direction; limited languages, fixed patterns, and a generic 9:16 safety model. |
Text adds the number, objection, or context voice cannot. It is not karaoke. Placement responds to face, product, UI, and device. |
Semantic captions + term review + face/product/UI collision tests in TikTok/Reels device previews. |
| Picture finish |
Scale/pad, LUTs, and size checks; no guaranteed semantic reframe, shot match, skin/exposure gate, or stabilization path. |
Corrects exposure/WB first, protects skin and product color, keyframes crop as attention moves, keeps useful phone texture. |
Tracked reframe + shot match with manual override and an explicit “keep for trust / remove as defect” decision. |
| Proof + CTA |
Can be requested in free text; not mandatory or verified in canonical cut artifacts. |
Shows a result, use, comparison, receipt, or real reaction while the claim is made; CTA is earned by the payoff. |
Proof ledger connecting every claim to a visible beat and checking CTA wording, visibility, and safe placement. |