See the shot
Subjects, scene, composition, palette, lighting, texture, and visible action.
Upload a short reference video. Get its camera language, motion, lighting, timing, audio—and production-ready prompts for the models you use.
Drop your video here
or click to browse your files
This original five-second reference clip was analyzed by ShotToPrompt. The excerpts below are from its generated result, not placeholder copy.
Locked side-profile tracking, low-key nocturnal light, a dark navy palette, warm orange subject, and continuous horizontal motion.
The rider pedals continuously while the camera follows at a fixed speed; soft rain ambience and subtle mechanical sounds support the motion.
“A minimalist 2D vector animation of an orange figure riding a white line-art bicycle across a dark street at night… The camera follows from a strict side-profile perspective in a smooth horizontal tracking shot.”
ShotToPrompt translates what the clip does into the production choices an AI video model can follow.
Subjects, scene, composition, palette, lighting, texture, and visible action.
Lens feel, camera position, motion path, transitions, pace, and temporal beats.
Editable prompts tailored to the strengths and conventions of each video model.
Your video is sent directly to Google Gemini for visual and audio analysis. ShotToPrompt does not save the uploaded video itself.
We keep the generated prompt, basic file metadata, and an anonymous usage identifier for a limited period so the tool can enforce its daily limit and improve reliability.
See exactly what we process and retainNo. A video does not contain its source prompt. ShotToPrompt analyzes observable production choices and creates a new, editable prompt that aims to reproduce its visual grammar.
MP4, MOV, WebM, and MPEG clips up to 15 seconds. The temporary beta upload limit is 4 MB. Short, visually coherent clips produce the most useful shot breakdowns.
No. The video is sent to Gemini for the analysis request and is not saved in the ShotToPrompt database. We retain the generated analysis and minimal operational metadata for limits and product quality.
Video models respond differently to motion, timing, audio, and structure. Each version preserves the same creative intent while emphasizing the controls that model handles best.