Learn
AI Tutor
News
Tools
Jobs
Glossary
Certification
Quizzes
Mission
Support
English
Search
⌘K
Submit AI Tool
Donate
English
Search
⌘K
Learn
AI Guides & Foundations
AI Tutor
Ask AI anything, free
News
Latest AI Developments
Tools
Top AI Directory
Jobs
AI Hiring Board
Glossary
AI Terms Dictionary
Certification
Get Your AI Certificate
Quizzes
Interactive AI Assessments
Mission
Why We Exist
Support
Help and Contact
Submit AI Tool
Donate
English
Home
/
Glossary
/
Vision-Language Model (VLM)
AI Glossary Term
What is Vision-Language Model (VLM)?
Definition
A multimodal model that jointly processes visual and textual information.
Related terms
Artificial Intelligence (AI)
The broad field of building systems that perform tasks requiring pattern recognition, reasoning, language, or decision-making.
CLIP
A multimodal model architecture that learns shared representations between text and images.
Computer Vision
The branch of AI that extracts meaning from images and video.
Context Window
The maximum amount of input tokens a language model can process at once.
Hallucination
When a model generates fluent but false or unsupported information.
Large Language Model (LLM)
A language model trained on massive text corpora to generate and analyze text.
Learn more in our free guides
AI Models Explained
Computer Vision
Model Collapse
Model Lifecycle
Browse the full AI Glossary
Explore all free AI guides