
OCDevel AI Video Generation Podcast
Writing a Video Prompt in Parts Instead of Wishes
A close look at why video prompts fail — broken down into seven controllable parts of a shot, how leaving any of them blank or overloading all of them wrecks a clip — followed by a quick tour of the current state of AI video tools and rankings. Episode page & show notes Visit website The seven slots of a video prompt A video prompt isn't a sentence, it's a form with seven fixed slots, and if you leave one blank the model fills it in with whatever's most average, which is where flat, stock-footage sludge comes from. The chapter walks through each slot in turn: subject (concrete, ruling-out detail rather than a category like "a woman"), action (one verb, one event, not a plot), camera move (static, push in, pull back, pan, orbit, handheld follow β defined plainly, with warnings about what happens when the slot is left blank), lens (wide vs. long, and what each does to background and space), lighting (direction, hardness, time of day), tone (mood and look kept separate, and kept to one or two words instead of a pile of adjectives), and pacing (how much can actually fit into a five-to-ten-second clip before motion starts to smear). It lays out a practical routine: write the seven labels down, fill each with a short phrase, render once cheaply, then change exactly one slot at a time and re-render to learn what your specific tool actually obeys. It also names the failure mode this same lesson tends to produce β an overstuffed prompt with multiple ideas crammed into every slot, which renders as generic porridge rather than as an error β and gives three signs to spot it (genericness, vague motion, small edits stop changing the output), plus the fix: strip back to three slots and rebuild. A related trap gets flagged too: a camera move that physically fights the action, like orbiting around a subject who's also running at the lens, which produces smearing and warped geometry rather than what was asked for. What moved this month A quick pass through the current state of video generation: a blind-comparison arena, read job by job rather than as one ranking, shows new entrants from Alibaba and a MiniMax variant pushing into the top of text-to-video, image-to-video and editing charts, while incumbents like Runway's Gen-4.5 have slid down the same table since spring. On silent (no-dialogue) boards, a model called HappyHorse leads outright. Release-wise, ByteDance's Seedance 2.5 now does thirty seconds in one pass with synchronized audio, dialogue and effects included at no extra charge. The overall pattern: audio in the same pass is becoming standard, clip lengths are stretching well past the old five-second ceiling, and price per second keeps dropping.


