A shot list with timecodes
Each cut gets its own start and end time, subject, action, scene, camera behavior, object movement, visible text, and sound notes.
Free · no login · no ad gate
Upload a short video and get editable prompts by shot, real start and end frames, sound notes, and formats for major AI video models.
Upload and analyze
Get prompts by shot, real start and end frames, and model-ready instructions in one analysis.
2. Upload your source video
Choose a short video to begin.
Report a problemYour source stays visible
Compare every generated shot description with the original video before you copy the prompt.
Direct answer
A video to prompt generator watches a reference clip and writes the instructions you need to rebuild it with an AI video model. Frame to Prompt detects cuts, describes each shot, extracts real frames from the uploaded file, and formats the result for the model you choose.
It does not pretend to recover the hidden original prompt. You get a reconstruction you can check: what was visible or audible, what remains uncertain, which frame to upload, and which text to paste for each shot.
Page reviewed September 15, 2026. Upload limits and output fields are read from the same product facts used by the live tool.
Concrete output
The result is built for the next action. You can compare it with the source, correct a mistaken detail, download the right frame, and copy a prompt without rewriting the whole analysis.
Each cut gets its own start and end time, subject, action, scene, camera behavior, object movement, visible text, and sound notes.
Download the opening, middle reference, and ending frame for each shot. The tool does not generate substitute images and label them as source frames.
If the model mistakes a pan for object movement or reads text incorrectly, edit that field once before copying any model-specific prompt.
Switch between Universal, Veo, Kling, Runway, Seedance, and Sora without analyzing or paying for the same video again.
From file to usable prompt
Keep the source short enough to review. A ten-second clip with three cuts is three generation tasks, even when it looks like one video in your editor.
Pick Veo when you want first-and-last-frame guidance, Runway when one starting image will carry the look, or Universal when you have not chosen a generator yet.
Use an MP4, MOV, or WebM file no longer than 60 seconds and strictly below 20 MiB (20,971,520 bytes). The original stays visible beside the result while you review it.
Play the source at each timecode. Correct the subject, physical action, framing, visible words, dialogue, music, or sound effect when the analysis is wrong.
Download the listed frame, paste only that shot prompt into your selected model, trim the result to the source timing, and place the clips in the stated order.
Choose the right output
A model prompt is part of a workflow, not a bag of descriptive words. The image inputs, supported duration, and treatment of sound change what belongs in the text box.
| Output | Frame to use | What the text should do |
|---|---|---|
| Universal | Opening, reference, and ending frames for review | Keep a complete visual and sound brief until you select a generation model. |
| Veo 3.1 | Opening and ending frames when that workflow is available | Describe one continuous transition, visual direction, dialogue, ambience, and sound effects. |
| Runway | Opening frame | Describe subject motion, environmental motion, camera motion, direction, speed, and timing in simple positive language. |
| Kling | Opening and ending frames when supported | Describe the physical transition between the two frames and keep subject identity consistent. |
Need the exact workflow? Open the Veo video prompt guide or the Runway video prompt guide.
Before you spend credits
A finished MP4 does not reveal which model created it, the original text prompt, a random seed, hidden reference images, camera hardware, or the editor's timeline. If another tool claims to recover those details exactly, the file itself does not provide enough evidence for that claim.
Frame to Prompt therefore keeps observed details inside the copyable prompt and puts inferred details in a separate review area. This reduces invented 4K, lens, lighting-equipment, and “masterpiece” language that may sound impressive but does not help a generation model reproduce the source.
The practical goal is a faster first attempt and a clearer correction loop. If the first result misses an action, change that action. If the composition drifts, use the extracted frame. If the text changes, rebuild it as an editor overlay.
Specific questions
Yes. The current public tool is free, requires no account or payment details, and has no ad gate. It allows up to three analyses per visitor each day, subject to a small site-wide daily capacity.
You can upload MP4, MOV, or WebM files no longer than 60 seconds and strictly below 20 MiB (20,971,520 bytes). If a MOV or WebM file will not preview in your browser, export an H.264 MP4 and try again.
No. A rendered video does not contain its original prompt, seed, model settings, or edit timeline. The tool reconstructs a usable prompt from visible and audible evidence, then keeps uncertain interpretations outside the copyable prompt.
A cut usually changes the subject, framing, location, or action. One prompt per shot lets you generate, replace, and trim each clip separately instead of asking one model call to reproduce several unrelated scenes.
Yes. Analyze the video once, then switch between Universal, Veo, Kling, Runway, Seedance, and Sora formats. The visual observations stay the same while frame instructions and prompt structure change for the selected workflow.
No. The prompt is a checked starting point. Exact identity, complex motion, readable text, timing, and sound can require more than one generation and a final edit. Review every shot beside the source before spending generation credits.