C.W.K.
Stream
Quiz · 4 questions

🏗️ Architectures: Vision, Sequence, and Attention

CNNs, recurrence, attention, transformers

Level 0Curious
0 XP0/73 lessons0/11 achievements
0/120 XP to next level120 XP to go0% complete

Quiz

01What's the most fundamental advantage of a CNN over an MLP for image classification?
02Why are residual connections so universal in modern architectures (ResNet, Transformer, U-Net, etc.)?
03What does scaled dot-product attention compute?
04What's the difference between encoder-only (BERT) and decoder-only (GPT) transformer attention?
Spotted a bug or have feedback on this page?Report an Issue

Comments 0

🔔 Reply notifications (sign in)
Sign inPlease sign in to comment.

No comments yet — be the first.