Skip to content

AI Scene Classification

The AI Scene Classification batch tool uses AI to describe your video thumbnails and add the results as keywords. You can access this tool from the start page or by right-clicking a video and selecting Classify Scenes... from the context menu.

The tool offers three options that can be used independently or together. They differ in what they recognise and how broad their vocabulary is:

Scene Classification -- the overall setting. It labels each thumbnail with one best-fit place category from a fixed list (e.g., indoor, kitchen, beach).

Object Detection -- specific objects that are present, from a fixed set of 80 common types (e.g., car, person, dog).

Image Tagging -- a broad, open-vocabulary description of the whole frame, spanning scenes, objects, settings and activities (e.g., beach, crowd, airplane, sunset). It is the most general option — it covers more than the fixed scene and object lists — and is downloaded on demand; see AI Image Tagging.

The minimum confidence slider applies to Scene Classification and Object Detection — it controls how certain the AI must be before adding a scene or object keyword; higher values reduce false positives but may miss some. Image Tagging is governed by its own Tag sensitivity instead (see its help page).

Select which videos to process: All videos, the currently selected videos, or the filtered set. You can also choose to skip thumbnails that have already been classified. The frame skip setting controls how many thumbnails to skip between processed ones, where 1 means all thumbnails are processed.

See the Scene Classification preferences for additional settings.

Object detection can also run with your own fine-tuned model to detect custom object types — see Custom Object Detection Models.