Foundation vision models trained on web-scale image data, transferable to many downstream computer-vision tasks.