For classification: Perform mutli-view testing that uniformly samples N clips from a video along its temporal axis. For each clip, it takes 3 crops to cover the spatial dimension, followed by averaging the softmax scores across all Nx3 views to form a video-level prediction. All
(test_loader, model, test_meter, cfg, writer=None)
source not stored for this graph (policy: none)
no test coverage detected