A base Vision Transformer model with 16x16 patch size and 384x384 input resolution for image classification.
No lifecycle dates recorded.
No resource links recorded.