Transformer
The neural network architecture behind essentially every modern language model, built around attention.
Introdotto nel 2017, il transformer ha sostituito l’elaborazione sequenziale con l’attenzione, permettendo al modello di guardare tutti i token dell’input insieme e di pesarne la rilevanza reciproca. È quel parallelismo ad aver reso pratico l’addestramento su testo su scala internet. La ‘T’ di GPT sta per transformer.
In pratica: In ‘il trofeo non entrava nella teca perché esso era troppo grande’, l’attenzione è il modo in cui il modello collega ’esso’ a ’trofeo’.
Where this comes up
- Claude Opus 4.7: The Most Powerful Opus Model Yet by Anthropic
- DeepSeek vs ChatGPT in 2026: Coding, Pricing, Privacy & Best Use Cases
- How Does ChatGPT Work? A Deep Dive into AI Language Models
- Sakana AI Fugu Review: Fugu Ultra vs Claude Fable 5
- The Best YouTube Channels to Learn AI: A Comprehensive Guide
- What Is Generative AI in Simple Terms?