Reinforcement Learning
Also called: RL
Learning by trial and error, guided by rewards rather than labelled answers.
Un agente di apprendimento per rinforzo agisce in un ambiente, osserva cosa succede e riceve un segnale di ricompensa. Nel corso di molti episodi impara una policy che massimizza la ricompensa attesa. L’RL sta dietro ai sistemi che giocano e, in forma modificata, dietro alla fase di allineamento dei moderni modelli di chat.
In pratica: Un agente impara a giocare senza che gli venga spiegata alcuna regola — solo un punteggio che sale o scende.
Where this comes up
- AI Learning Roadmap for Beginners: Your Comprehensive Guide
- AI Training for Logistics Coordinators: Your Path to Success
- AI for Trading Course: Your Complete Guide to Learning and Success
- Best AI Side Hustles in 2026: 50 Beginner Ideas, Tools & Pay
- Claude vs ChatGPT 2026: Which AI Is Better for You?
- Grok vs ChatGPT in 2026: Benchmarks, Pricing & Best Use Cases