Models / intel / DPT SwinV2 Large 384

DPT SwinV2 Large 384

JSON →
vision

A monocular depth estimation model using a Swin Transformer V2 large backbone with 384x384 input resolution, developed by Intel.

imagevision
Specs
Context & limits
context window
max output
Lifecycle
Dates
released2023-12-01
Resources