BERT
AIBERT: Google's 2018 encoder-only model, trained by masking words, which set the pattern of pretraining then fine-tuning for language tasks.
BERT: Google's 2018 encoder-only model, trained by masking words, which set the pattern of pretraining then fine-tuning for language tasks.