Models / google / ViT Base Patch16 224

ViT Base Patch16 224

JSON →
googlevision

A Vision Transformer (ViT) base model pre-trained on ImageNet-21k and fine-tuned on ImageNet-1k at 224x224 resolution.

imagevision
Specs
Context & limits
context window
max output
Lifecycle
Dates

No lifecycle dates recorded.

Resources