Impara
Tutor IA
Notizie
Strumenti
Lavori
Glossario
Certificato
Quiz
Missione
Supporto
English
Search
⌘K
Invia strumento
Dona
English
Search
⌘K
Impara
AI Guides & Foundations
Tutor IA
Ask AI anything, free
Notizie
Latest AI Developments
Strumenti
Top AI Directory
Lavori
AI Hiring Board
Glossario
AI Terms Dictionary
Certificato
Get Your AI Certificate
Quiz
Interactive AI Assessments
Missione
Why We Exist
Supporto
Help and Contact
Invia strumento
Dona
English
Home
/
Glossary
/
Inference Endpoint
AI Glossary Term
What is Inference Endpoint?
Definition
A deployed API interface that receives model requests and returns predictions in production.
Related terms
API (Application Programming Interface)
A structured way for one software system to send requests to and receive responses from another system.
Inference
The runtime phase where a trained model generates predictions or outputs.
Usage-Based Billing
Pricing where costs scale with API calls, tokens, inference time, or consumed compute.
Inference-Time Compute
The amount of processing power consumed while producing each response.
Decision Tree
A model that makes predictions through a sequence of if-then feature splits.
Ensemble
Combining predictions from multiple models to improve robustness or accuracy.
Learn more in our free guides
AI Inference
AI Inference Optimization
LLM Inference Routing and Load Balancing
Seldon Core and Inference Graphs
Browse the full AI Glossary
Explore all free AI guides