Knowledge & Data · Model asset

Contrastive Image-Text Encoder

Model assetKnowledge & DataKnowledge & Dataarc:ContrastiveImageTextEncoder

A pair of text and image encoders trained jointly with contrastive loss so matching text-image pairs map to nearby vectors in one shared space.

Responsibility. Maps text and images into a shared embedding space.

deployed onis optimized byInference Server: deployed onInference ServerEngine Builder: is optimized byEngine Builder
Direct neighbourhood (hover for relationship types)

Relationships

deployed on structural

is optimized by lifecycle

Quantitative guidance

As stated by the sources; verify before use.

Classification

Patterns
Contrastive learningZero-shot classification
Technologies
CLIPOpenCLIPNV Embed

Sources

  1. Ch2.7: T. Nguyen, "Multimodal RAG Approaches," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 2.7. ISBN: 9798244538229.