Навчання
ШІ-репетитор
Новини
Інструм.
Вакансії
Глосарій
Сертифікат
Тести
Місія
Підтримка
English
Search
⌘K
Додати інстр.
Донат
English
Search
⌘K
Навчання
AI Guides & Foundations
ШІ-репетитор
Ask AI anything, free
Новини
Latest AI Developments
Інструм.
Top AI Directory
Вакансії
AI Hiring Board
Глосарій
AI Terms Dictionary
Сертифікат
Get Your AI Certificate
Тести
Interactive AI Assessments
Місія
Why We Exist
Підтримка
Help and Contact
Додати інстр.
Донат
English
Home
/
Glossary
/
Reinforcement Learning
AI Glossary Term
What is Reinforcement Learning?
Definition
Training by reward signals where an agent learns actions that maximize long-term return.
Related terms
Reinforcement Learning from Human Feedback (RLHF)
A training method that uses human preference signals to shape model behavior.
Reward Model
A model that scores outputs based on preference signals, often used in RLHF pipelines.
AI Agent
A software system that can observe, reason, and take actions to achieve a goal, often using tools and memory.
Constitutional AI
A training and behavior-shaping approach where model outputs are guided by a fixed set of written principles.
Agentic Loop
An iterative cycle where an AI agent observes, plans, acts, and reflects until it completes a goal or hits a stop condition.
DPO (Direct Preference Optimization)
A training method that fine-tunes models directly on preference pairs without needing a separate reward model.
Learn more in our free guides
Deep Learning
Reinforcement Learning
Machine Learning Basics
Supervised Learning
Browse the full AI Glossary
Explore all free AI guides