vision foundation models

These are large pre-trained models designed for various vision tasks (e.g., object detection, image segmentation) that serve as a baseline for transfer learning or fine-tuning on specific applications, often leveraging vast quantities of unlabeled data for training.

8 papers