In self-attention, queries and keys undergo a dot product matrix multiplication to produce a score matrix that determines how much focus should be placed on each word relative to other words, with higher scores indicating more focus.
factualpending
Speaker
Unidentified Speaker — Illustrated Guide to Transformers Neural Network: A step by… [4Bdc55j80l8]Evidence Quote
“the queries and keys undergoes a dot product matrix multiplication to produce a score matrix the score matrix determines how much focus should a word be put on other words so each word will have a score to correspond to other words”
Created: 8/13/2026, 9:51:15 AM
My Notes
Loading notes...