cs.CVSep 28, 2026

Revisiting Risky Tackle Detection with Vision Transformers

Authors: Syed Ahsan Masud Zaidi, Lior Shamir, Scott Dietrich

Organizations: Kansas State University, Manhattan, KS, USA · Albright College, Reading, PA, USA

Abstract

This paper is a Track 2 reproducibility companion to an ICPR 2026 study on risky tackle detection in American football prac- tice videos. The original work fine-tuned a Video Vision Transformer (ViViT) on 733 clips labeled with the SATT-3 rubric. It used focal loss, Taguchi L18 augmentation, and 5-fold cross-validation. It reported risky- class recall of 0.67 and risky-class F1 of 0.59. This companion documents the released artifact and traces those numbers to specific scripts, fold out- puts, and aggregation files. The reproduced headline is run_15. It com- bines Gaussian noise with static brightness decrease and uses no rotation and no flip. Its fold-mean risky recall is 0.667 and its fold-mean risky F1 is 0.588. These values match the published headline after rounding. The ablation shows that brightness is the dominant factor. Its risky-recall main-effect range is 0.055, which is larger than the ranges for rotation, flip, and noise. Without augmentation, ViViT reaches risky recall of 0.545 and does not exceed the C3D baseline of 0.583. The raw clips show iden- tifiable student athletes, so they cannot be redistributed. The artifact provides a public sample for pipeline checks and a controlled route for full-data review.

Figures & tables

Explore similar work

CardsList
  1. CheckOne: Lightweight Fault Detection and Mitigation for Vision Transformers

    Aug 3, 2026Mohammad Hasan Ahmadilivani, Sven-Markus Loorits, Jaan RaikSelf-Supervised Vision TransformersVision Transformer

  2. EviViT: Evidence-Adaptive Vision Transformers for Fine-Grained Perception

    Sep 29, 2026Yaoxin Niu, Zhangquan Chen, Yang Zhang +5Self-Supervised Vision TransformersFine-Grained Perception

  3. An Open-Source Two-Stage Computer Vision Pipeline for Fine-Grained Vehicle Classification using Vision Transformers

    Jun 3, 2026Gandhimathi Padmanaban, Fred FengAutomotive DetectionComputer Vision