by DeepMind
A few-shot visual-language model that processes interleaved sequences of images and text.
No reviews yet. Be the first.
No problems reported for this model.