Transformer Architecture Quiz
Quiz covering Transformers and Attention
Transformer Architecture Quiz
5 questions | Pass: 70% | Earn 25 points
Questions in this quiz
A preview of the 5 questions covered. Start the quiz above to answer them, check your score, and read the explanations.
- 1
What is the primary purpose of the 'Self-Attention' mechanism in a Transformer model?
- 2
In the original Transformer architecture, why are 'Positional Encodings' added to the input embeddings?
- 3
What is the function of the 'Feed-Forward Network' (FFN) that follows the Multi-Head Attention layer in each Transformer block?
- 4
When training a Transformer, what is the role of the 'Mask' in the decoder's masked self-attention layer?
- 5
How does increasing the number of 'Attention Heads' in Multi-Head Attention affect the model's learning capacity?
Enjoying the courses?
Everything stays free. Pro shows fewer ads, doubles the points you earn on every lesson and quiz so you progress twice as fast, unlocks half of every practice exam — plus full case studies — with the Learn & Exam study modes, and lets you read each lesson on one page.
- ✓ Fewer advertisements
- ✓ 2× points per lesson & quiz
- ✓ 50% of every exam unlocked
- ✓ Learn & Exam modes
- ✓ Distraction-free lessons