Learn
AI Tutor
News
Tools
Jobs
Glossary
Certification
Quizzes
Mission
Support
English
Search
⌘K
Submit AI Tool
Donate
English
Search
⌘K
Learn
AI Guides & Foundations
AI Tutor
Ask AI anything, free
News
Latest AI Developments
Tools
Top AI Directory
Jobs
AI Hiring Board
Glossary
AI Terms Dictionary
Certification
Get Your AI Certificate
Quizzes
Interactive AI Assessments
Mission
Why We Exist
Support
Help and Contact
Submit AI Tool
Donate
English
Home
/
Glossary
/
Model Quantization
AI Glossary Term
What is Model Quantization?
Definition
Reducing numeric precision of model weights to decrease memory and inference cost.
Related terms
Quantization
Converting model weights to lower precision formats such as 8-bit or 4-bit.
QLoRA
A fine-tuning technique that combines 4-bit weight quantization with LoRA adapters to reduce memory needs.
Flash Attention
An optimized attention algorithm that reduces memory use and speeds up transformer training and inference.
Memory (Agent Memory)
Stored context an AI agent uses across steps or sessions to improve continuity.
AI Agent
A software system that can observe, reason, and take actions to achieve a goal, often using tools and memory.
AI Safety
A field focused on reducing harmful behavior, failures, and misuse risks in AI systems.
Learn more in our free guides
AI Models Explained
Model Collapse
Model Lifecycle
Model Context Protocol
Browse the full AI Glossary
Explore all free AI guides