For stock video creators

Stock video keyword generator with multi-frame analysis

We read three frames in time order from every clip and merge the results, and measure the camera move itself across the whole clip — so your keywords capture scene changes plus real moves like Zoom In, Panning and Tracking Shot that single-frame tools miss.

No credit card required — start with 15 free credits.

Tagging video like a photo loses the part buyers search for: the motion. A single frame can't tell whether the camera zoomed, panned, or tracked a subject — so single-frame tools miss camera-movement keywords entirely. PixTagger samples three frames from each clip in time order (start, middle, end), and merges the keywords into one list. Camera movement is measured separately — we track how the picture actually moves across the whole clip rather than inferring it from a few stills, so a slow gimbal drift is caught and a locked-off tripod shot stays silent. It also adds the keywords the platforms require for video search, and a clip costs 2 credits no matter how long it runs.

See it in action

Watch the workflow, start to finish.

Three frames, time-ordered

We sample the start, middle, and end of every clip and compare them — catching scene changes and action a single frame can't show.

Camera-movement detection

When the movement is clear we add the exact controlled-vocabulary term — Zoom In, Zoom Out, Panning, Tilt, Tracking Shot, or Aerial View — and stay silent when it's ambiguous, because a wrong movement keyword is worse than none.

Adobe + Getty video CSV

Includes the 'video' keyword Adobe Stock requires for video search, flags vertical clips, and exports the CSV format each platform expects.

Two credits per clip, any length

No per-second pricing. A 5-second loop and a 60-second clip each cost the same 2 credits.

How it works

From shoot to upload-ready in under a minute.

1

Switch to Videos mode & drop clips

Frames are decoded locally in your browser, so the full-resolution file never has to upload.

2

Add Shoot Notes (optional)

Context sharpens the merged keywords across all three frames.

3

Generate

Vision AI runs on each frame and merges everything into one ordered keyword list; the camera move is measured from the clip itself.

4

Review & download CSV

Edit inline, then export the Adobe or Getty video CSV — import-ready.

FAQ

Questions, answered.

How does multi-frame video tagging work?+

For every clip we extract three representative frames (start, middle, end), run vision AI on each, and merge the keywords, which catches scene changes a single frame misses. Camera movement is a separate optional step: we measure how the picture moved across the clip in your browser, so the move that reaches your CSV is one we actually observed.

Which video formats are supported?+

MP4, MOV, and M4V — the formats Adobe Stock and Getty accept. Longer clips are sampled at start, middle, and end.

Does it detect camera movement?+

Yes, when it's clear: Zoom In, Zoom Out, Panning, Tilt, Tracking Shot, and Aerial View, using each platform's exact term. If the movement is ambiguous we add no movement keyword rather than guess.

How many credits does a video cost?+

2 credits per clip, regardless of length.

Does camera-movement measurement work in every browser?+

Standard MP4 / H.264 is measured in any browser. ProRes, HEVC and cinema-grade H.264 have to be decoded by the browser to be measured, and Chrome and Firefox skip them — tag those clips in Safari if you want movement keywords, or leave the box unticked and name the move in your shoot notes, which always wins over the measurement anyway. Tagging itself works everywhere regardless; only the movement keyword depends on this.

My clip has complex action — are three frames enough?+

Usually yes, and for a short single-action clip more frames add nothing. But for a long take that travels through a scene, an activity that turns into another, or a reveal at the end, tick "Complex action" and we sample five points instead of three, so the merged keywords describe the whole sequence rather than the parts three frames happened to land on. That costs 1 extra credit per clip.

How many clips can I tag at once?+

30 per drop — lower than the 100-photo cap because we decode frames from every clip in your browser before anything is sent, which is heavier on the tab than resizing a photo. The cap is per drop, not per export: keep dropping batches and they stack in the same review screen and download as one combined CSV, and re-dropping a file already in the export is skipped rather than charged again.

Tag your footage the way buyers search for it.

Try PixTagger free. No credit card. 15 images on the house.

Get started free