Skip to content

model

CLIP

Contrastive language-image pretraining model whose paired text and image encoders are widely reused to align natural-language descriptions with visual features.

Current clusters