We are loading the service guidance for this workflow.
We are loading the dataset structure and key columns for this workflow.
filename
target
path
video_name
label
tag
Check the file for missing information, unsuitable columns, and other issues that may affect the results.
Resolve blocking issues, review warnings, then choose models and start training.
These models train today using extracted frame and clip features. They are the canonical video workflow currently wired for training, evaluation, and prediction.
These architectures are the next canonical direction for video, analogous to the deep backbones used in image and audio services. The first native 3D CNN and video-transformer backbones can now be enabled when runtime dependencies are available on this machine.
Your recommended setup appears after the data check is complete.
AutoML will compare a practical mix of suitable models. Change it only when you have a specific reason.
Your estimate appears after you choose a setup.