Apprendre
Tuteur IA
Actus
Outils
Emplois
Glossaire
Certificat
Quiz
Mission
Aide
English
Search
⌘K
Ajouter outil
Donner
English
Search
⌘K
Apprendre
AI Guides & Foundations
Tuteur IA
Ask AI anything, free
Actus
Latest AI Developments
Outils
Top AI Directory
Emplois
AI Hiring Board
Glossaire
AI Terms Dictionary
Certificat
Get Your AI Certificate
Quiz
Interactive AI Assessments
Mission
Why We Exist
Aide
Help and Contact
Ajouter outil
Donner
English
Home
/
Glossary
/
Inference Endpoint
AI Glossary Term
What is Inference Endpoint?
Definition
A deployed API interface that receives model requests and returns predictions in production.
Related terms
API (Application Programming Interface)
A structured way for one software system to send requests to and receive responses from another system.
Inference
The runtime phase where a trained model generates predictions or outputs.
Usage-Based Billing
Pricing where costs scale with API calls, tokens, inference time, or consumed compute.
Inference-Time Compute
The amount of processing power consumed while producing each response.
Decision Tree
A model that makes predictions through a sequence of if-then feature splits.
Ensemble
Combining predictions from multiple models to improve robustness or accuracy.
Learn more in our free guides
AI Inference
AI Inference Optimization
LLM Inference Routing and Load Balancing
Seldon Core and Inference Graphs
Browse the full AI Glossary
Explore all free AI guides