Skip to content
Learn
News
Tools
Jobs
Mission
Search
⌘K
English
Donate
Menu
×
Start here
Start learning
Learn
Go to Learn
Courses
Guides
AI Tutor
Prompt library
Quizzes
Certification
Glossary
News
Go to News
Latest AI news
Topic trackers
Research & datasets
Blog
Editorial standards
Corrections
Tools
Go to Tools
Tool directory
Best-of lists
Compare tools
Cost calculator
Prompt Refiner
Submit AI Tool
Jobs
Go to Jobs
Browse AI jobs
Post a job
Sponsor AIU
Mission
Go to Mission
Mission
Impact & stats
Authors
Help centre
What's new
Support
Donate
Home
/
Glossary
/
Distilled Model
AI Glossary Term
What does “
Distilled Model
” mean?
Definition
A smaller model trained to imitate a larger model's behavior while using less compute at inference.
Related terms
Inference-Time Compute
The amount of processing power consumed while producing each response.
Knowledge Distillation
Training a smaller model to imitate the outputs of a larger model.
Inference
The runtime phase where a trained model generates predictions or outputs.
Pruning
Removing less important model weights or neurons to reduce size and compute.
Test-Time Compute
Additional inference computation used during response generation to improve quality or reasoning.
Usage-Based Billing
Pricing where costs scale with API calls, tokens, inference time, or consumed compute.
Learn more in our free guides
AI Models Explained
Model Collapse
Model Lifecycle
Model Context Protocol
See also
Differential Privacy
Embedding Model
Data Provenance
Eval Harness
Data Lineage
Feature Store
Constitutional AI
Groundedness
Previous
Differential Privacy
Next
Embedding Model
Browse the full AI Glossary
Explore all free AI guides