Models / hustvl / YOLOS Base

YOLOS Base

JSON →
vision

A vision transformer-based object detection model that treats detection as a sequence-to-sequence problem, similar to DETR but using a pure transformer encoder-decoder architecture.

imageobject-detection
Specs
Context & limits
context window
max output
Lifecycle
Dates
released2021-06-01
Resources