Jifunze
Mwalimu wa AI
Habari
Zana
Kazi
Kamusi
Cheti
Majaribio
Dhamira
Usaidizi
English
Search
⌘K
Wasilisha zana ya AI
Changia
English
Search
⌘K
Jifunze
Miongozo ya AI na misingi
Mwalimu wa AI
Ask AI anything, free
Habari
Maendeleo mapya ya AI
Zana
Orodha bora ya AI
Kazi
Bodi ya ajira za AI
Kamusi
Kamusi ya istilahi za AI
Cheti
Pata cheti chako cha AI
Majaribio
Tathmini shirikishi za AI
Dhamira
Kwa nini tupo
Usaidizi
Usaidizi na mawasiliano
Wasilisha zana ya AI
Changia
English
Home
/
Glossary
/
Model Quantization
AI Glossary Term
What is Model Quantization?
Definition
Reducing numeric precision of model weights to decrease memory and inference cost.
Related terms
Quantization
Converting model weights to lower precision formats such as 8-bit or 4-bit.
QLoRA
A fine-tuning technique that combines 4-bit weight quantization with LoRA adapters to reduce memory needs.
Flash Attention
An optimized attention algorithm that reduces memory use and speeds up transformer training and inference.
Memory (Agent Memory)
Stored context an AI agent uses across steps or sessions to improve continuity.
AI Agent
A software system that can observe, reason, and take actions to achieve a goal, often using tools and memory.
AI Safety
A field focused on reducing harmful behavior, failures, and misuse risks in AI systems.
Learn more in our free guides
AI Models Explained
Model Collapse
Model Lifecycle
Model Context Protocol
Browse the full AI Glossary
Explore all free AI guides