|
|
|
|
|
by brcmthrowaway
2 days ago
|
|
> The core language model has no mechanism for representing its prompt as opposed to any other part of its current input sequence; indeed it has no mechanism for cross-reference from one part of the sequence to another. (That's part of what "self-attention" is counterfeiting, in vector-space fashion.) -- Cosma Shalizi |
|