I recorded one raw intro clip, pointing at where I wanted things to go. GPT-6 Astra and HyperFrames turned it into a finished edit: a title, cards, captions, the websites I've built sliding past and last week's video playing in a window. In this video I break down how it works, then run it live. Here's the five-step process: Step 1, Transcribe: It transcribes your raw clip word by word. You can use OpenAI's Whisper model or an external tool like ElevenLabs, which is a bit faster. Step 2, Cut: From the transcript it makes the cuts: silences, stumbles, the dead space at the start and the end. Step 3, Plan the beats: Every cue you say becomes a moment on screen. When I say the cards coming in from the left, cards come in from the left. If you record yourself pointing and saying it, most of the plan is already done. Step 4, Build with HyperFrames: HyperFrames is HeyGen's locally hosted video editor. Astra pushes the cut video and the beat plan into it and builds the motion layers on top. Step 5, Verify: It watches the result back. If something's wrong it replans the beats and rebuilds. In the live run I show the exact prompt: where the raw clip is, which cues I say, where my websites and last video live. I ask for a plan first. It comes back as a timed map of every cue and what goes there. I say go, it builds the edit. When I wanted one change, keeping my background when I shrink into the corner, I just asked for it in the chat. To get started you need the HyperFrames skill. Copy the GitHub link, paste it into your GPT-6 Astra chat and tell it to install the skill. At the end I show three shorts built with HyperFrames from my long-form videos. Fair warning: Astra burns through tokens. For the quality you get, it's worth it. Resources: HyperFrames → https://github.com/heygen-com/hyperframes