Skip to content

Topic

Vision transformers

Image models built on transformer attention over patch embeddings rather than convolutions, widely used as pretrained backbones.

Current clusters