Skip to main content

6 docs tagged with "neural-networks"

View all tags

Attention Mechanism

Understand content addressing, Q/K/V, masks, multi-head variants, efficient implementations, and interpretation limits through a numerical example.

Multilayer Perceptron

Understand MLP capacity, backpropagation, optimization failures, and inductive bias through tensor shapes and a worked XOR construction.