CLIPVisionModel

E1312484 UNEXPLORED

CLIPVisionModel is a vision transformer-based image encoder from OpenAI's CLIP framework that maps images into a joint multimodal embedding space for tasks like image-text matching and retrieval.

All labels observed (1)

Label Occurrences
CLIPVisionModel canonical 1

How this entity was disambiguated

Referenced by (1)

Full triples — surface form annotated when it differs from this entity's canonical label.