Reinforcement Learning
Also called: RL
Learning by trial and error, guided by rewards rather than labelled answers.
A reinforcement learning agent acts in an environment, observes what happens, and receives a reward signal. Over many episodes it learns a policy that maximises expected reward. RL is behind game-playing systems and, in modified form, behind the alignment step in modern chat models.
In practice: An agent learns to play a game with no rules explained — only a score that goes up or down.
Where this comes up
- AI Learning Roadmap for Beginners: Your Comprehensive Guide
- AI Training for Logistics Coordinators: Your Path to Success
- AI for Trading Course: Your Complete Guide to Learning and Success
- Best AI Side Hustles in 2026: 50 Beginner Ideas, Tools & Pay
- Claude vs ChatGPT 2026: Which AI Is Better for You?
- Grok vs ChatGPT in 2026: Benchmarks, Pricing & Best Use Cases