Models / depth-anything / Prompt Depth Anything ViT-L

Prompt Depth Anything ViT-L

JSON →
vision

A promptable depth estimation model using a Vision Transformer large backbone, enabling conditional depth prediction from user prompts.

imagevision
Specs
Context & limits
context window
max output
Lifecycle
Dates
released2024-11-01
Resources