vision transformers

A type of transformer model specifically designed for computer vision tasks. They adapt the attention mechanism of transformers to process images by dividing them into patches, leading to improvements in accuracy and computational efficiency over traditional convolutional neural networks.

37 papers