Large Language Models, End Term
January 2024 term, 28 Apr 2024, Set QDB1
BERT stands for
Bidirectional Encoder Representation for Text
Bidirectional Encoder Representation from Transformer
Bidirectionaly Extracted Representation from Transformer
Bidirectionaly Extracted Representation of Text
Sign in to report a problem with this question.
You cannot change your answers after submitting.
The palette shows the status of every question. Pick a number to go straight to it.
BERT stands for The statement that, in general, position information can also be injected into the attention layers of a transformer model instead of the embedding of input tokens is Figure from the original question paper