Skip to content
Learn
News
Tools
Jobs
Mission
Search
⌘K
English
Donate
Menu
×
Start here
Start learning
Learn
Go to Learn
Courses
Guides
AI Tutor
Prompt library
Quizzes
Certification
Glossary
News
Go to News
Latest AI news
Topic trackers
Research & datasets
Blog
Editorial standards
Corrections
Tools
Go to Tools
Tool directory
Best-of lists
Compare tools
Cost calculator
Prompt Refiner
Submit AI Tool
Jobs
Go to Jobs
Browse AI jobs
Post a job
Sponsor AIU
Mission
Go to Mission
Mission
Impact & stats
Authors
Help centre
What's new
Support
Donate
Home
/
Glossary
/
CLIP
AI Glossary Term
What does “
CLIP
” mean?
Definition
A multimodal model architecture that learns shared representations between text and images.
Related terms
Diffusion Model
A generative architecture that learns to reverse noise to synthesize images, audio, or other content.
Convolutional Neural Network (CNN)
A neural architecture optimized for processing grid-like data such as images.
Embedding
A numeric vector representation that captures semantic meaning of text, images, or other data.
Generative AI
AI systems that produce new content such as text, images, audio, video, or code.
Multimodal Model
A model that can process or generate multiple data types such as text, image, and audio.
OCR (Optical Character Recognition)
Technology that converts text in images or scans into machine-readable text.
Learn more in our free guides
Gradient Clipping
CLIP and Vision-Language Models
See also
Classifier
Compute
Classification
Computer Vision
Chain-of-Thought
Context Window
Calibration
Continual Learning
Previous
Classifier
Next
Compute
Browse the full AI Glossary
Explore all free AI guides