Enables machines to interpret and understand visual information such as images and video.
3 models
by Meta AI
A promptable zero-shot image segmentation model trained on 1.1B masks over 11M images.
Not yet ratedby Google Research
Applies transformers directly to image patches for classification; inspired many CV variants.
Not yet ratedby Ultralytics
The latest YOLO family model for real-time object detection, classification, and segmentation.
Not yet rated